How we sample prompts and pages
We start from the demand side. For a category we build a prompt frame of 300–600 buying-intent prompts covering product discovery, comparison, budget bands, use case, and retailer-led phrasing. Prompts are run across the assistants that matter for the category and the cited sources are recorded per answer.
The cited URLs, not the prompts, are the unit of analysis. We rank cited domains and pages by citation frequency, then crawl the pages that carry the most citation weight — typically roundups, reviews, and category hubs on editorial publishers. Every crawled page is retained with its fetch timestamp so a result can be reproduced against the version we saw.
