AI citation tracking is the practice of recording which sources an AI answer links to — and how that changes over time, across prompts and between competitors.
It answers a different question from mention tracking. A mention tells you whether your brand is named. A citation tells you which pages the answer relied on or pointed to. Either can happen without the other.
#Why citations matter
- Citations show where answer evidence comes from. If three review sites and one comparison article account for most links in your category, those are the pages shaping the answer.
- They show who is trusted by the engine for a topic — and whether that's you, a competitor or a third party.
- They give you a concrete work list: pages to improve, sources to earn, claims to correct.
#What to record
For every answer, store:
| Field | Why it matters |
|---|---|
| Prompt, stage and wording | Results are only comparable for the same prompt |
| Engine and model version | Different engines cite differently |
| Country and language | Sources vary by market |
| Date and run number | Answers vary run to run |
| Cited URLs, in order | Order and presence both carry information |
| Anchor or passage context | Shows what the source supports |
| Whether the brand was named | Keep mentions separate from citations |
Store answers as collected. Rewriting or summarising them loses the evidence you'll want when someone challenges a number.
#Classify sources
A flat list of 400 domains isn't useful. Classify each domain so patterns appear:
- Owned — your verified domains.
- Competitor — rivals' domains.
- Editorial — publications and independent guides.
- Review and comparison sites.
- Community — forums, Q&A, social.
- Documentation and reference.
- Marketplaces and directories.
- Video.
Then ask which types each engine leans on for each stage of the journey.
#Metrics
- Owned-citation rate — answers that link to your domains ÷ valid answers.
- Source share — a domain's share of all citations in a prompt set.
- Gap sources — domains cited alongside competitors but not for you.
- Corroboration — whether a claim about you appears on independent sources as well as your own.
Report the denominator every time, and keep citation rates separate from mention rates.
#Turning findings into action
- Fix access first. If your own pages can't be fetched, nothing else matters. See AI crawler access.
- Make your pages the best source for a specific question. State the answer plainly, add evidence and keep it current.
- Close gaps legitimately. For sources that cite rivals but not you: is there a real reason to be listed or reviewed? A better source to offer? A factual correction to request?
- Avoid manufactured endorsements. Fake comparison sites and paid pseudo-reviews mislead readers and tend to backfire.
#Common mistakes
- Treating the order of citations as a ranking. It may reflect how the answer was composed, not source quality.
- Counting a citation to a competitor's page about you as an owned citation.
- Ignoring variance. A source that appears in one run of ten is different from one that appears in nine.
- Comparing different prompt sets between periods.
#Tooling options
You can start with a spreadsheet and a handful of prompts. Once you need repeat runs, competitors and history, a purpose-built tool saves time. Source Map in Citeroot (learn more) classifies domains and produces gap lists from the answers it collects.
#Sources
- OpenAI, Overview of OpenAI crawlers.
- Google Search Central, Optimizing for generative AI features in Search.
- Aggarwal et al., GEO: Generative Engine Optimization.