Sources & Citations

Find the pages AI reads before it answers

Answer engines summarise sources, they do not invent recommendations. See every domain and URL cited in your category, how each is classified, and which ones your competitors are winning without you.

For most teams this is the fastest route to a visibility gain.

Sourcesretrievals in range

Retrievals

4,234

↑ 12%
G
g2.comListicle1,284
R
reddit.comDiscussion963
northlane.comYou741
C
capterra.comComparison528
E
everline.comCompetitor402

Every

Citation stored per run

3 levels

Domain, host and URL

19

Source types classified

1 click

From a gap to a brief

Domains

Every source behind every answer

An engine builds its answer by retrieving pages and summarising them. We store every one of those citations, for every run, rolled up by domain with retrievals, citation rate and which way each is trending.

  • Retrievals over time, plotted per scan run rather than per calendar day
  • Domain movers, so you see influence shifting before it costs you
  • Bookmarks on the domains your team is actively working on
Sourcesretrievals in range

Retrievals

4,234

↑ 12%
G
g2.comListicle1,284
R
reddit.comDiscussion963
northlane.comYou741
C
capterra.comComparison528
E
everline.comCompetitor402

Classification

Know what kind of source you are dealing with

Every domain and URL is typed. A listicle, a review comparison, a Reddit thread and your own docs all get cited, but the work required to win each one is completely different.

  • Types include listicle, comparison, discussion, editorial, profile and product
  • Your own pages, competitor pages and third parties separated out
  • Type mix tracked over time, so you can see the category shifting
Source distributiondomain types
Listicle
28%
Discussion / UGC
22%
Comparison
16%
Editorial
13%
You
12%
Competitor
9%

Gap analysis

The pages that recommend your competitors

The most valuable report we produce. It finds the sources cited in answers where competitors appear and you do not, then ranks them by retrievals multiplied by how many competitors benefit.

  • Split by domain, host and individual URL so you know exactly what to target
  • Scored, so the top row is genuinely the top row
  • Hands straight off to the content engine as a scoped brief
Gap analysiscompetitors cited, you are not

g2.com/categories/ecommerce-analytics

Score 94
ListicleUsed by 4 competitors312 retrievals

reddit.com/r/ecommerce

Score 81
DiscussionUsed by 3 competitors244 retrievals

capterra.com/ecommerce-analytics

Score 73
ComparisonUsed by 3 competitors187 retrievals

shopify.com/partners/directory

Score 58
ProfileUsed by 2 competitors121 retrievals

Site audit

Check the engines can read you at all

Before any of the above matters, an AI crawler has to be able to fetch and parse your site. The audit runs live HTTP checks and tells you which of them you are failing.

  • Homepage reachable
  • robots.txt present
  • AI crawlers allowed
  • llms.txt present
  • Structured data (JSON-LD)
  • Sitemap available
  • Homepage indexable
Site auditAI search health

Checks passed

5/7

2 to fix
Homepage reachablePass
robots.txt presentPass
AI crawlers allowedPass
Homepage indexablePass
Sitemap availablePass
llms.txt presentFail
Structured data (JSON-LD)Fail

Also included

Built to answer the question underneath the number

Not just which sites get cited, but which ones are worth your next two weeks.

URL level drill-in

Go from a domain to an individual page, and from that page to every answer it was ever cited in and every brand mentioned on it.

Hosts kept separate

Subdomains are tracked as their own hosts and rolled up separately, because docs.example.com and example.com behave nothing alike.

Bookmarks

Flag domains and URLs at project level so everyone is working from the same target list rather than their own notes.

Source trends

Watch a domain's influence rise or fall across runs. Engines change what they trust and you want to notice early.

Cited pages you own

See which of your own pages the engines actually read. It is frequently not the ones you optimised.

API access

Pull the full citation dataset into your own warehouse or reporting stack through the project API.

The handoff

A gap is only useful if someone can act on it

Source gaps become scored opportunities on the content backlog, with the evidence already attached, so nobody redoes the research.

Content opportunitiesrefreshed each scan

GA4 vs dedicated ecommerce analytics

Losing promptScore 94
Generate

Cheapest analytics tool for small stores

Weak promptScore 81
Generate

Get listed on G2 ecommerce category

Source gapScore 76
Generate

Integrations coverage

Topic holeScore 64
Generate

See which sites decide your category

Run a scan and get your source gap report. Most teams find at least one directory they should have been on for years.