Sources & Citations
Find the pages AI reads before it answers
Answer engines summarise sources, they do not invent recommendations. See every domain and URL cited in your category, how each is classified, and which ones your competitors are winning without you.
For most teams this is the fastest route to a visibility gain.
Retrievals
4,234
Every
Citation stored per run
3 levels
Domain, host and URL
19
Source types classified
1 click
From a gap to a brief
Domains
Every source behind every answer
An engine builds its answer by retrieving pages and summarising them. We store every one of those citations, for every run, rolled up by domain with retrievals, citation rate and which way each is trending.
- Retrievals over time, plotted per scan run rather than per calendar day
- Domain movers, so you see influence shifting before it costs you
- Bookmarks on the domains your team is actively working on
Retrievals
4,234
Classification
Know what kind of source you are dealing with
Every domain and URL is typed. A listicle, a review comparison, a Reddit thread and your own docs all get cited, but the work required to win each one is completely different.
- Types include listicle, comparison, discussion, editorial, profile and product
- Your own pages, competitor pages and third parties separated out
- Type mix tracked over time, so you can see the category shifting
Gap analysis
The pages that recommend your competitors
The most valuable report we produce. It finds the sources cited in answers where competitors appear and you do not, then ranks them by retrievals multiplied by how many competitors benefit.
- Split by domain, host and individual URL so you know exactly what to target
- Scored, so the top row is genuinely the top row
- Hands straight off to the content engine as a scoped brief
g2.com/categories/ecommerce-analytics
reddit.com/r/ecommerce
capterra.com/ecommerce-analytics
shopify.com/partners/directory
Site audit
Check the engines can read you at all
Before any of the above matters, an AI crawler has to be able to fetch and parse your site. The audit runs live HTTP checks and tells you which of them you are failing.
- Homepage reachable
- robots.txt present
- AI crawlers allowed
- llms.txt present
- Structured data (JSON-LD)
- Sitemap available
- Homepage indexable
Checks passed
5/7
Also included
Built to answer the question underneath the number
Not just which sites get cited, but which ones are worth your next two weeks.
URL level drill-in
Go from a domain to an individual page, and from that page to every answer it was ever cited in and every brand mentioned on it.
Hosts kept separate
Subdomains are tracked as their own hosts and rolled up separately, because docs.example.com and example.com behave nothing alike.
Bookmarks
Flag domains and URLs at project level so everyone is working from the same target list rather than their own notes.
Source trends
Watch a domain's influence rise or fall across runs. Engines change what they trust and you want to notice early.
Cited pages you own
See which of your own pages the engines actually read. It is frequently not the ones you optimised.
API access
Pull the full citation dataset into your own warehouse or reporting stack through the project API.
The handoff
A gap is only useful if someone can act on it
Source gaps become scored opportunities on the content backlog, with the evidence already attached, so nobody redoes the research.
GA4 vs dedicated ecommerce analytics
Cheapest analytics tool for small stores
Get listed on G2 ecommerce category
Integrations coverage
See which sites decide your category
Run a scan and get your source gap report. Most teams find at least one directory they should have been on for years.