[{"data":1,"prerenderedAt":137},["ShallowReactive",2],{"product-/en/product/page-audit-en":3},{"id":4,"title":5,"body":6,"description":6,"extension":7,"meta":8,"navigation":94,"path":131,"seo":132,"stem":135,"__hash__":136},"content_en/1.product/page-audit.yml","Page Audit",null,"yml",{"category":9,"hero":10,"takeaway":20,"features":24,"technical":44,"benefits":72,"comparison":88,"faq":105,"cta":121},"Features",{"eyebrow":11,"headline":12,"description":13,"links":14},"Be Retrieved, the third rung","Audit a page against [published thresholds]{class='text-primary'}","Most content scores are a letter grade and a vague note to write better. This one names the band your page fell outside, and by how much.",[15],{"label":16,"to":17,"size":18,"color":19},"14-day free trial","https://app.trustdata.tech","lg","primary",{"label":21,"text":22,"attribution":23},"Key takeaway","Heading density below **0.5 per 100 words** leaves nothing to chunk on. Above **3.0**, headings stop meaning anything. Both ends fail, which is the whole argument for bands over a score: \"write better\" becomes a number that is out of range.","Every band has a ceiling as well as a floor",{"title":25,"description":26,"items":27},"What gets checked","Every page is fetched server-side and parsed, so the audit sees what a crawler sees rather than what a browser renders.",[28,32,36,40],{"title":29,"description":30,"icon":31},"13 binary signals","Get a list of what is missing, not a grade. Heading hierarchy, FAQ block, lists, structured data, a key takeaway, images, external links, statistics, an intro summary under 80 words, a table, a named author, freshness and authoritative sources, each present or absent with nowhere to hide.","i-lucide-list-checks",{"title":33,"description":34,"icon":35},"11 structural bands","Argue with the thresholds instead of accepting a verdict. Eleven bands each with a floor and a ceiling, paragraphs **30 to 90 words**, list ratio 0.10 to 0.40, heading density 0.5 to 3.0 per 100 words and eight more, all published as numbers.","i-lucide-ruler",{"title":37,"description":38,"icon":39},"Signals for what your site is","Get audited as the business you actually are. A lead-generation site is additionally checked for comparison content, testimonials, case studies and social proof; an ecommerce site for product schema, visible pricing and reviews.","i-lucide-store",{"title":41,"description":42,"icon":43},"Technical state alongside","Stop reconciling the audit against a separate crawl. Status code, indexability, robots directive, canonical, title, meta, H1 count, schema types and link counts come from the same pass.","i-lucide-wrench",{"title":45,"description":46,"features":47},"Why bands rather than a score","A composite grade tells you a page is weak. A band tells you which number to change.",[48,52,56,60,64,68],{"title":49,"description":50,"icon":51},"Macro, document architecture","Heading depth, section count, heading density and whether a table of contents exists. This is the level that decides whether an engine can find the passage that answers a question, rather than the page as a whole.","i-lucide-layout-template",{"title":53,"description":54,"icon":55},"Meso, information chunking","Average and 90th-percentile paragraph length, list ratio and chunk variance. Uneven chunking is the failure mode a word count cannot see: a page can average 60 words a paragraph and still hide a 400-word wall.","i-lucide-align-left",{"title":57,"description":58,"icon":59},"Micro, visual emphasis","Emphasis density, blockquotes, code density and image-to-text ratio. Under-emphasised text gives an extractor nothing to anchor on; over-emphasised text gives it too much and the signal flattens.","i-lucide-highlighter",{"title":61,"description":62,"icon":63},"Bands, not thresholds to maximise","Each band has a floor and a ceiling because both ends fail. Too few headings and there is no structure; too many and every heading means less. The bands come from published research on generative-engine extraction and are re-tunable against observed citation rates.","i-lucide-git-compare-arrows",{"title":65,"description":66,"icon":67},"A readable diff on every change","Each crawl stores a snapshot. When a page changes, the audit reports which fields moved: title, meta description, H1, structured data, body copy. So a score change is always attributable to an edit.","i-lucide-history",{"title":69,"description":70,"icon":71},"Rolls up to the rung","The average page audit score across your audited pages is the Be Retrieved rung on the readiness ladder. One page's failure is a task; the average moving is a programme working.","i-lucide-chart-no-axes-column",{"title":73,"description":74,"items":75},"What it changes","The point of a numeric band is that it ends the argument about whether a page is good enough.",[76,80,84],{"title":77,"description":78,"icon":79},"It separates writing from structure","Stop telling a writer to write better when structure is the fault. Most pages that fail extraction are well written and badly structured, and a 400-word paragraph with no headings is not a quality problem.","i-lucide-scissors",{"title":81,"description":82,"icon":83},"It makes a rewrite checkable before publishing","Check a draft before you ship it. The bands are computable on unpublished content, so you never have to publish, wait a fortnight and infer from citation counts whether structure was the problem.","i-lucide-file-check",{"title":85,"description":86,"icon":87},"It feeds a measured experiment","Prove the rewrite worked. A page change becomes a change event with a field-level diff, and the prompts it should have moved become the treatment group against a control drawn from those it should not.","i-lucide-flask-conical",{"title":89,"description":90,"items":91},"Against a content score","The difference is whether the output names a number you can change.",[92,96,99,101,103],{"feature":93,"trustdata":94,"ga4":95,"other":95},"Published numeric bands with floors and ceilings",true,false,{"feature":97,"trustdata":94,"ga4":95,"other":98},"Server-side fetch, sees what a crawler sees","Sometimes",{"feature":100,"trustdata":94,"ga4":95,"other":95},"Signals differ by site type",{"feature":102,"trustdata":94,"ga4":95,"other":95},"Field-level diff between crawls",{"feature":104,"trustdata":94,"ga4":95,"other":95},"Rolls up into a readiness rung",[106,109,112,115,118],{"label":107,"content":108,"defaultOpen":94},"What is an AI content extractability audit?","It scores whether a page can be parsed, chunked and quoted by an AI engine, which is a different question from whether it ranks. TrustData fetches each page server-side, parses the HTML, and checks 13 binary signals along with 11 structural bands covering document architecture, information chunking and visual emphasis. The output is the specific bands the page fell outside, not a grade.",{"label":110,"content":111},"Why are the structural checks bands rather than targets?","Because both ends fail. Heading density below 0.5 per 100 words means the page has no structure to chunk on; above 3.0 means headings stop carrying meaning. Paragraphs under 30 words fragment an idea across chunks; over 90 they bury the answer. A list ratio under 0.10 means nothing is enumerable, over 0.40 means the prose has become a bullet dump with no connective reasoning. Every band has a floor and a ceiling for that reason.",{"label":113,"content":114},"Where do the thresholds come from?","They derive from published experimental work on generative-engine extraction, and they are deliberately treated as a first pass rather than a law. The intention is to regress them against observed citation rates on real properties, so the bands tighten as evidence accumulates. Until then they are documented rather than hidden, which is the difference between a threshold you can argue with and a black-box score you cannot.",{"label":116,"content":117},"Does the audit differ by site type?","Yes. Beyond the shared signals, a lead-generation site is checked for comparison content, testimonials, case studies, use cases and social proof, because those are what an engine looks for when answering an evaluation question about a service. An e-commerce site is checked for product schema, visible pricing and reviews instead. Scoring both against the same list would penalise one for lacking something it should not have.",{"label":119,"content":120},"How does this connect to the rest of the ladder?","The average audit score across your pages is the Be Retrieved rung, which is third. It sits above Get Found, because a page a crawler cannot reach cannot be extracted whatever its structure, and below Be Trusted, because a perfectly structured page still needs corroboration to be cited with confidence. If Get Found is failing, the audit is not your bottleneck yet, and the ladder will say so by marking the lower rung as the focus.",{"title":122,"description":123,"links":124},"Audit a page against the bands","14-day free trial. Page audits are part of AI Visibility, included in every plan.",[125,127],{"label":16,"to":17,"size":126,"color":19},"xl",{"label":128,"to":129,"variant":130,"size":126},"See the full readiness ladder","/en/product/ai-visibility","outline","/product/page-audit",{"title":133,"description":134},"AI content extractability audit, in bands","Score any page against the structural signals AI engines use to extract and cite content. Named bands with numbers, not a vague content grade.","1.product/page-audit","NXjacoxsjBy9L0O29MRFq9LNN-4fJCcMn41FwT8kTGs",1786740930232]