[{"data":1,"prerenderedAt":139},["ShallowReactive",2],{"product-/en/product/prompt-monitoring-en":3},{"id":4,"title":5,"body":6,"description":6,"extension":7,"meta":8,"navigation":94,"path":133,"seo":134,"stem":137,"__hash__":138},"content_en/1.product/prompt-monitoring.yml","Prompt Monitoring",null,"yml",{"category":9,"hero":10,"takeaway":20,"features":24,"technical":44,"benefits":72,"comparison":88,"faq":107,"cta":123},"Features",{"eyebrow":11,"headline":12,"description":13,"links":14},"Be Chosen, the fifth rung","Track the prompts that [hold still]{class='text-primary'}","AI answers vary between runs. A prompt whose brand list changes every probe measures noise, not visibility. Each carries a stability score so you know which is which.",[15],{"label":16,"to":17,"size":18,"color":19},"14-day free trial","https://app.trustdata.tech","lg","primary",{"label":21,"text":22,"attribution":23},"Key takeaway","`reproducibility.py` computes the **median pairwise Jaccard similarity** across a prompt's recent brand lists. Below the floor, the prompt is flagged rather than charted, because a line drawn through noise looks exactly like a line drawn through a result.","An unstable prompt is noise, not a trend",{"title":25,"description":26,"items":27},"Prompts you did not have to invent","Guessing at prompts produces a list that flatters the brand. These come from your audience and from real demand.",[28,32,36,40],{"title":29,"description":30,"icon":31},"Generated from your ICP","Stop writing the prompts you already win. Define your audience once and prompts are generated from those personas and from observed search demand, then suggested in bulk for you to edit or reject.","i-lucide-user-check",{"title":33,"description":34,"icon":35},"Four question types","Fix the buying stage you are losing. Brand direct, category, comparison and pain point map to different stages, fail for different reasons and are scored separately rather than averaged into one number.","i-lucide-message-square",{"title":37,"description":38,"icon":39},"A reproducibility score per prompt","Know whether a trend is real before you act on it. The score is the median pairwise similarity of the brand lists returned across recent probes, so low means the answer is unstable and any trend drawn from it is noise.","i-lucide-repeat",{"title":41,"description":42,"icon":43},"A quality score per prompt","See which prompts are worth keeping. **0 to 1** across four axes: how well it discriminates you from competitors, how stable its mention rate is, whether it connects to a decision anyone acted on, and whether it has been probed recently.","i-lucide-gauge",{"title":45,"description":46,"features":47},"How a prompt is run and read","Every number on a prompt traces to the probes behind it, and the responses are kept.",[48,52,56,60,64,68],{"title":49,"description":50,"icon":51},"Two modes on the same schedule","Each prompt runs with live web search and without. The gap between the two separates what the model knows from what it just read, which is the difference between a brand problem and a content problem.","i-lucide-split",{"title":53,"description":54,"icon":55},"Full response retained","The answer behind every result is stored and readable. A mention rate you cannot open is a number you have to take on faith.","i-lucide-file-text",{"title":57,"description":58,"icon":59},"History per prompt","Mention rate, position and sentiment over time for each individual prompt, so a fall in the aggregate can be traced to the specific questions that moved.","i-lucide-trending-up",{"title":61,"description":62,"icon":63},"Probe budget and prompt limits","50 monitored prompts on Measure through 400 on Prove, against monthly probe budgets of 3,000 to 24,000. The scheduler caps at the budget rather than billing you past it.","i-lucide-gauge-circle",{"title":65,"description":66,"icon":67},"Bulk actions and priority","Prompts can be prioritised, paused and edited in bulk, because a prompt set is a living list rather than a one-time setup step.","i-lucide-list-checks",{"title":69,"description":70,"icon":71},"Feeds Be Chosen","Brand mention rate across the prompt set scores rung 5,with the AI Overview gap count surfaced alongside it as evidence.","i-lucide-chart-no-axes-column",{"title":73,"description":74,"items":75},"Why reproducibility is on the page","It is the metric that decides whether the rest of the numbers mean anything.",[76,80,84],{"title":77,"description":78,"icon":79},"It marks which trends are real","Avoid presenting a convincing chart of nothing. A prompt with a low **reproducibility score** moves persuasively and means nothing, and publishing that score beside the trend is the difference between measurement and decoration.","i-lucide-activity",{"title":81,"description":82,"icon":83},"It tests our own claim","Hold us to our own claim. We say persona-grounded prompts return more stable brand lists than invented ones, and reproducibility is the metric that proves or disproves it on your data, published either way.","i-lucide-flask-conical",{"title":85,"description":86,"icon":87},"It stops you optimising against noise","Spend the quarter on a question that will hold still. Chasing a mention-rate dip on an unstable prompt costs a sprint and produces nothing, because the dip was never there.","i-lucide-shield-check",{"title":89,"description":90,"items":91},"Against a prompt list you wrote yourself","The difference is where the prompts come from and whether their stability is known.",[92,97,99,101,104],{"feature":93,"trustdata":94,"ga4":95,"other":96},"Prompts generated from ICP and real demand",true,false,"Manual entry",{"feature":98,"trustdata":94,"ga4":95,"other":95},"Reproducibility score per prompt",{"feature":100,"trustdata":94,"ga4":95,"other":95},"Quality score across four axes",{"feature":102,"trustdata":94,"ga4":95,"other":103},"With-search and no-search on the same schedule","Rarely",{"feature":105,"trustdata":94,"ga4":95,"other":106},"Full response retained and readable","Sometimes",[108,111,114,117,120],{"label":109,"content":110,"defaultOpen":94},"What is AI prompt tracking?","It is the practice of running a fixed set of questions against AI engines on a schedule and recording how your brand appears in the answers. TrustData runs each prompt in two modes, with live web search and without, across four question types, and keeps the full response behind every result. Prompt counts run from 50 on Measure to 400 on Prove, against monthly probe budgets of 3,000 to 24,000.",{"label":112,"content":113},"What is a reproducibility score and why does it matter?","It is the median pairwise similarity between the brand lists an engine returned across a prompt's recent probes. If an engine returns roughly the same set of brands each time, the prompt is stable and a trend drawn from it means something. If the set changes every run, the prompt is unstable and its chart is noise that will still look like a trend. Publishing the score next to the trend is what stops a team spending a quarter chasing a movement that was never real.",{"label":115,"content":116},"Where do the prompts come from?","From your own audience definition and from observed search demand, rather than from a brainstorm. You describe your ideal customer profile once, prompts are generated from those personas and from real query data, and they are suggested in bulk for you to edit, keep or reject. The reason is straightforward: a prompt list written from inside a company tends to ask the questions that company already wins, which produces a flattering baseline and no information.",{"label":118,"content":119},"What does the quality score measure?","Four axes on a 0 to 1 scale. Discriminance, weighted highest, is how far your mention rate diverges from your competitors' on that prompt: a question everyone wins tells you nothing. Stability is the variation in its daily mention rate. Actionability is whether the prompt is connected to a recommendation anyone actually followed. Recency is whether it has been probed in the last fortnight. Below three extracted probes the score is withheld rather than estimated.",{"label":121,"content":122},"Why run each prompt with and without web search?","Because they answer different questions. With search enabled, the engine retrieves current pages, so the result reflects your content and your crawlability. With search disabled, it answers from training data, so the result reflects what the model already believes about you. Strong with search and weak without means you are being retrieved but are not established in the model. The reverse means you are established but not being retrieved, which is usually a robots or extractability problem.",{"title":124,"description":125,"links":126},"Monitor the questions your buyers actually ask","14-day free trial. Prompt monitoring is part of AI Visibility, included in every plan.",[127,129],{"label":16,"to":17,"size":128,"color":19},"xl",{"label":130,"to":131,"variant":132,"size":128},"See the full readiness ladder","/en/product/ai-visibility","outline","/product/prompt-monitoring",{"title":135,"description":136},"AI prompt tracking with a stability score","Monitor the questions your buyers actually ask, generated from your ICP and search demand, with a stability metric on every prompt.","1.product/prompt-monitoring","VD72eqTX0LDqeLlGheW7CZ_O1D3gjioFcGc2_3s4N7Q",1786740930420]