1. Get Found
Find out whether AI crawlers can reach you at all. Half your money pages blocking AI bots caps this rung at weak however much else is crawled, and no content work moves it until you unblock them.
Key takeaway
Readiness is expressed as the identity of the lowest-scoring rung, not an average. A single 0 to 100 number would let a strong content score hide a blocked crawler, and the two failures need opposite fixes.
The readiness ladder
AI readiness is not one number. A brand can be crawled thoroughly and still never be cited, or be cited constantly and never referred. Each rung is scored 0 to 100 from its own signal, with 60 as the healthy cut-off. Exactly one rung is marked the focus rung: the lowest one below 60 whose prerequisites are already clear. Rungs above a broken rung render as locked, because earning them first does not work.
Find out whether AI crawlers can reach you at all. Half your money pages blocking AI bots caps this rung at weak however much else is crawled, and no content work moves it until you unblock them.
Check that machines resolve you to the right entity. Where no entity profile exists yet, schema coverage stands in, and knowledge graph gaps surface as evidence for your next fix rather than as a worse score.
See whether your content survives extraction, scored across 13 binary signals and 11 structural bands. Your content team gets named failures, heading depth or paragraph length, instead of a grade to interpret.
Read the framing around your brand, scored on the positive share of every mention. A high mention rate with negative sentiment is worse than absence, and this rung tells you which one you have.
See whether you get named when the question is competitive, measured as share of voice with your AI Overview gaps listed as evidence. That gap list is the brief for your next content sprint.
Watch how AI referred traffic behaves, as sessions and conversion rate, labelled emerging because they are proxies. We do not claim a citation caused a sale, so this rung tells you where to look rather than what to bank.

AI Overviews
AI Overviews are the most volatile surface anyone tracks. They are generated, personalised and A/B bucketed, so a single weekly reading cannot carry the change detection people ask of it. TrustData reports presence as a rolling rate with the observation count behind it, so a keyword read once and a keyword read twenty-eight times are never confused.
Stop guessing whether AI Overviews affect you. Every tracked keyword carries how often one actually appeared over four weeks, with the observation count behind it, so a thin sample shows its count rather than a misleading rate.
See where you sat among the sources when an Overview appeared, and the snippet Google used. Position 7 of 8 is not the win a presence flag suggested it was.
Track the number that moves when your content changes: how often you were cited on the occasions an Overview did appear. Presence is the market, citation is your share of it.
Read the trade on one row. AI Overview state sits beside impressions, clicks, position and conversions, so your SEO lead sees a blue link lost against a citation won.

AI Mode
Google now answers on two AI surfaces and most tools report them as one. AI Overviews sit above the blue links and are read per prompt. AI Mode is a separate conversational surface with its own answer, its own references and its own cited domains, and it is read once per tracked keyword. That difference is why AI Overview tracking is bundled in every tier while AI Mode is priced per tracked keyword, from 29 EUR.
Check AI Mode separately, because it writes its own answer instead of summarising the result page. Holding the AI Overview citation on a keyword can still leave you absent here, which is why the two are never averaged.
See which domains Google chose and where you sat among them. This is the competitor view of the surface, not a presence flag on your own brand.
Add AI Mode without touching your probe budget. It is read once for each keyword you track, so the cost follows your keyword list and there is no per-probe overage.
Compare both Google surfaces on the same keyword row. A single blended Google score hides the keyword where you won the Overview and lost AI Mode.

Engines and probes
Every probe runs in two modes, with live web search and without. The gap between them separates what the model knows from what it just read, which is the difference between a brand problem and a content problem.
Pay for the surfaces your buyers actually use. ChatGPT, Perplexity and Google AI Overview ship with every tier including the entry tier; Gemini, Claude, Mistral and DeepSeek are add-ons from 15 EUR a month, and Google AI Mode is priced per tracked keyword from 29 EUR.
Read the gap between with-search and no-search to get your diagnosis. Strong with search and weak without means you are in the index but not the model; weak both ways means the content is not extractable at all.
Score brand direct, category, comparison and pain point separately. Each maps to a different buying stage and fails for a different reason, so averaging them would hide the one you can fix this quarter.
Use the 0 to 100 composite as your supporting number, not your headline. It returns nothing below 5 probes rather than reporting noise, and the ladder is what tells you where to act.

Every rung is computed from a surface you can open, drill into and export. Nothing is a black box score.
Server-side logs of which AI crawler operators fetched which pages, and which pages your robots.txt is turning away. Feeds Get Found.
Extractability scored against 13 binary signals plus structural bands for heading depth, section count, paragraph length, list ratio and emphasis density. Feeds Be Retrieved.
Schema markup coverage and validity per page, plus the entity corroboration that decides whether machines resolve you to the right thing. Feeds Be Understood.
Presence rate, your position, citation rate and the shown snippet, per keyword, over a rolling four-week window with its observation count.
Which sources get cited alongside you, and which of your own pages earn citations. The gap between the two is your outreach list.
Every mention classified positive, neutral or negative, aggregated over 14 days. Feeds Be Trusted.
Named competitors tracked in every probe, with share of voice broken down per engine and per question type.
The prompt set you monitor, with history per prompt and the full response behind every result. 50 prompts on Measure, up to 400 on Prove.
Prompts generated from your own audience definition and real search demand, not guessed. Suggested in bulk, then edited by you.
Which communities AI engines lean on when answering about your category, so you know where the answer is actually being formed.
Every change becomes a measured experiment against a control group drawn from the prompts it should not have touched. Verdict after 7 days on an adjusted delta.
Every content change recorded with a field-level diff: title, meta description, H1, structured data, body copy. When something moves, you can see what you changed.
Two regimes of evidence, never bridged. This is the part most GEO tools leave vague.
Read leading indicators as leading indicators. Citation rate, share of voice and ranking lift always ship with an explicit noise level, stated on the page rather than in a footnote.
Get a control group without building one. When a change targets comparison queries your brand and category prompts keep running untouched, and the verdict is the adjusted delta, which filters out model updates and competitor moves.
Never take a citation to the CFO as a sale. Rung 6 reports sessions and conversion rate as proxies, labelled emerging, because causal claims belong to paid media where holdout tests can carry them.
Crawl access, extractability, citations, sentiment and prompt coverage, each measured on its own terms rather than averaged into a score.
Rung 1. Which engines fetched which page, read from your own logs, and whether a citation arrived with no crawl behind it.
Rung 3. Structural bands with a floor and a ceiling, published as numbers rather than folded into a letter grade.
Rung 4. Owned and earned kept apart rather than averaged, and the co-cited list read as an outreach brief.
Rung 4. How you are described when you are mentioned, with linked and unlinked mentions counted separately because they move for different reasons.
Rung 5. The questions your buyers actually ask, and a stability metric that can disprove our own prompt set.
FAQ
AI visibility monitoring tracks how your brand appears in AI-generated answers across engines like ChatGPT, Perplexity, Gemini and Claude, and in Google AI Overviews. Over 70% of discovery queries now return a zero-click AI answer: category research, comparisons, recommendations. Where your brand appears in those answers, or fails to, directly affects traffic. TrustData probes ChatGPT and Perplexity daily on every plan, reads AI Overviews per keyword, and scores readiness on six independent rungs rather than one blended index.
Because an average hides the one thing you can act on. A brand can score well on crawling and still be invisible in answers, or be mentioned constantly and never referred. Averaging six signals into one number tells you that something is wrong without telling you what. TrustData scores each rung independently against a healthy cut-off of 60, then marks exactly one focus rung: the lowest rung below the cut-off whose prerequisites are already clear. Rungs above an unhealthy rung are shown as locked, because working on them before the lower rung is fixed does not produce a result. A rung with too little data reads as insufficient data and is excluded from focus selection rather than scored as zero.
Get Found (can AI crawlers reach and retrieve you), Be Understood (do machines resolve you to the right entity), Be Retrieved (does your content survive extraction), Be Trusted (is the framing around your brand positive), Be Chosen (do you get named on competitive questions), and Enable Transactions (does AI referred traffic behave). They are scored in that order because each depends on the ones below it. There is no point optimising sentiment on pages AI crawlers cannot reach.
AI Overview state is read per tracked keyword and reported as a rolling four-week presence rate with the observation count behind it, not as a yes or no. When an Overview appears we also record whether you were referenced, your position among the sources, and the snippet shown, then compute a citation rate over the same window. Presence tells you the surface exists for that query; citation rate tells you your share of it. All of it sits on the same keyword row as impressions, clicks, position and conversions, so you can see the trade between a lost blue link and a won citation. AI Overviews are generated, personalised and A/B bucketed, which is exactly why a single reading rendered as a boolean overstates what anyone actually knows.
TrustData probes 6 engines: ChatGPT, Perplexity, Gemini, Claude, Mistral and DeepSeek, and reads two Google SERP surfaces, AI Overviews and AI Mode. ChatGPT and Perplexity ship as baseline probe engines on every plan, alongside AI Overview tracking. Gemini (25 to 199 EUR), Claude (99 to 799 EUR), Mistral and DeepSeek (15 to 99 EUR each) are per-engine, per-tier monthly add-ons. Google AI Mode is the exception: it is read once per tracked keyword rather than probed, so it is priced per keyword at 29 to 299 EUR. There is no per-probe overage: engine cost is amortised into the add-on price and the scheduler caps at your tier's probe budget. Grok was dropped in July 2026 rather than left on the list, because it never returned a usable row.
The Brand Health Score is a 0 to 100 composite measuring how consistently your brand appears in AI answers. It is a weighted average of your with-search mention rate (60%) and no-search mention rate (40%) across all engines and question types, updated daily. It returns no value at all below 5 probes rather than reporting a number built on noise. It is a supporting metric. The readiness ladder, not this score, is how you decide what to do next.
When you follow a recommendation, TrustData records the change with a field-level diff and starts tracking the prompts it should have affected. Those are the treatment group. Your other prompts keep running unchanged on the same schedule, and their drift is the natural control. The verdict is the adjusted delta, treatment change minus control change, which filters out noise from model updates, competitor activity and seasonality. A winner verdict requires a 10 point adjusted delta after at least 7 days, a deliberately high bar given typical probe sample sizes.
We do not claim that, and we will not report it as if we could. TrustData runs two regimes of evidence and never bridges them. Paid media, audiences and smart links are measured causally through geo, time and platform holdout tests, reported with explicit confidence intervals. AI visibility, content and SEO are measured as leading indicators: citation rate, share of voice, ranking lift, each with an explicit noise level. The sixth rung reports AI Search sessions and their conversion rate as proxy signals, labelled emerging. Any tool telling you it has attributed revenue to a ChatGPT citation is overstating what the data can carry.
They are two different Google surfaces. An AI Overview is the generated summary that appears above the blue links on a normal results page, and TrustData reads it per prompt as a rolling four-week presence rate. AI Mode is a separate conversational surface with its own answer, its own references and its own cited domains, and TrustData reads it once per tracked keyword. Because they are read differently they are billed differently: AI Overview tracking is included in every tier, while AI Mode is priced per tracked keyword from 29 EUR. They are reported side by side and never blended into one Google score, because a brand can hold the AI Overview citation on a keyword and be absent from AI Mode on the same keyword.
14-day free trial
14-day free trial. AI visibility is included in every plan, with ChatGPT, Perplexity and AI Overview tracking on every tier.