About

An independent product and content exploration.

Sourceframe exists to think carefully in public about a hard problem: making the public web usable, verifiable, and current for AI systems.

Most writing about web data for AI is either vendor marketing or a shallow tutorial. This project tries for something else: a consistent, technically literate treatment of discovery, retrieval, normalization, extraction, monitoring, and citation — with the trade-offs and failure modes left in.

Everything here is written as design work. Where a claim would need evidence to be responsible, the claim is framed as a question or an open experiment instead. Where a number would need measurement, none is given.

A design document should be judged by the failure modes it names, not by the confidence of its diagrams.
Working principle — editorial standard

What this is

  • A conceptual product design for a web-data platform aimed at AI systems.
  • A body of technical writing: field guides, workflow patterns, and research notes.
  • A set of browser-based planning aids that help structure a decision.
  • A record of design experiments, including the ones that have not been run yet.

What this is not

  • An operating company, a commercial service, or a product you can buy.
  • A live API. Nothing on this site retrieves, crawls, or stores web pages.
  • A source of benchmark results, uptime figures, or customer outcomes.
  • An endorsement of, or comparison against, any real vendor.

Working principles

01Current sources

Context is only useful if it reflects the page as it exists now.

02Structured outputs

Defined shapes beat free text when software has to consume the result.

03Traceable context

Every claim keeps a path back to the URL it came from.

04Production-minded workflows

Designs that assume failure, drift, and rate limits from the start.