Fixed-price audit · Two weeks · Booked one brand at a time

The Agentic Shelf Audit: How AI Models See Your Brand

Your brand now has a shelf position inside AI answers. The Agentic Shelf Audit is an AI brand visibility and model-belief audit that tests how major assistants describe your company, judge your claims, compare you with competitors and decide whether to recommend you — then traces those answers back to the sources and information gaps shaping them.

Every day, assistants answer questions about your ingredients, your health claims, your controversies — and decide whether you're the answer to your category's problems. They do it with confidence, at scale, from sources nobody in your building has audited. The Agentic Shelf Audit shows you exactly what models believe, where every belief comes from, and how to change it.

Get a free 3-belief teaser See scope & pricing

€7,500 fixed · One brand, one market · Delivered in 10 working days

The two questions

Nobody asks an AI to "recommend your brand." It doesn't matter.

Buyers use assistants in two ways that decide your revenue anyway: they ask for solutions to problems — and a brand is the answer. Or they name your brand and ask for a verdict — and the model delivers one. The audit measures both.

Question one · Presence

Are you in the answer?

"Best breakfast for a diabetic." "What helps atopic skin in winter." "Which formula for a colicky baby." Nobody names a brand — but every answer contains brands. We measure your Share of Answer: how often you appear at your category's entry points, how favorably, and who owns the default answer instead.

Question two · Belief

Is what they say about you true?

"Is that immunity claim real?" "Does it contain palm oil?" "Is it ultra-processed?" Models judge your brand daily — synthesizing studies, regulator decisions, forums, and stale product data into confident verdicts. We inventory and grade every belief in the Model Belief Report.

Why now

Models don't just recommend. They adjudicate.

One verdict

Search gave you ten links and let the buyer decide. An assistant hands down a single confident judgment — on your claims, your ingredients, your controversies — and the buyer takes it as read. The agent is now the user, and it doesn't read marketing.

Unaudited sources

Every model belief has a supply chain: retailer product data, Wikipedia, health portals, forums, old press. Nobody at your company has ever audited it — which means nobody is managing it.

Compounding

Beliefs models hold today feed the answers they give for years, and what they learn now hardens with every training cycle. Wrong beliefs are cheapest to fix while they're still forming.

Scope · Six lenses

What the audit actually inspects

A structured battery of real buyer queries — problem, interrogation, and comparison — run across the major assistants and agents, then traced back through the machine side of your presence to explain why the answers come out the way they do.

01 · Share of Answer

Are you the answer to your category's problems?

Problem and occasion queries built from your category's real entry points, run across leading assistants. We measure where you appear, how favorably, in whose company — and which competitors own the default answer.

02 · Belief inventory

What do models assert about you?

Every factual and evaluative belief models hold about your brand, extracted into a graded ledger — each weighted by severity and how entrenched it is across models:

03 · Claim adjudication

Do your claims survive the model's judgment?

Your full claim register — on-pack, on-site, in-campaign — tested claim by claim against model verdicts: upheld, hedged, rejected, or reversed onto a competitor. Including where models over-claim on your behalf, a legal exposure most brands have never seen.

04 · Comparability

How do you survive a head-to-head?

When an agent builds a comparison, someone decided which attributes matter. We test agent-constructed comparisons against named competitors and private label — and identify the missing proof points that cost you the verdict.

05 · Actionability

Can an agent actually buy from you?

The full agentic path on your owned channels and key retail listings: find, parse, basket, checkout. Every stall, misread, and abandonment is a documented revenue leak — often rerouted straight to a competitor. This is the commerce edge of agentic experience design.

06 · Source graph

Where does every belief come from?

Priority beliefs traced to their sources — retailer data, knowledge bases, press, forums, your own inconsistent pages — each graded for authority, accuracy, and whether you can fix it. This is the map that turns findings into influence.

Deliverables

What lands on your desk

Deliverable AShare of Answer Scorecard

Your presence and favorability across the category's entry points, benchmarked against up to three competitors and private label — with the full evidence pack and a re-runnable query battery, yours to keep.

Deliverable BThe Model Belief Report

The graded belief ledger, the claim-by-claim adjudication, the controversy narratives, and a threat register ranking every damaging belief by severity, spread, and confidence — with your Belief Integrity Score as the number leadership tracks.

See a redacted sample excerpt →

Deliverable CThe Source Graph

Every priority belief traced to the sources feeding it, graded by authority and addressability. Single points of failure, canon vacuums, and what a fix changes next week versus next training cycle.

Deliverable DInfluence roadmap + executive readout

Every fix ranked by impact and addressability — owned, influenceable, earned — cut into now / next / sustained. Delivered in a 60-minute session built for brand, e-commerce, legal, and comms in the same room.

How it runs

Ten working days. Two hours of your team's time.

Days 1–2

Intake. A 45-minute call plus your claim register and substantiation file. We lock the market, entry points, competitor set, and query battery — then you go back to your day job.

Days 3–6

Fieldwork. The full battery across assistants and agents — problem, interrogation, and comparison queries, plus agentic purchase attempts. Every transcript logged as evidence.

Days 7–8

Grading & tracing. Beliefs extracted, graded, and ranked; priority beliefs traced to their sources. Nothing enters the report without evidence and reproduction steps behind it.

Days 9–10

Synthesis & readout. Scorecard, Belief Report, Source Graph, and roadmap assembled — then delivered live to your leadership. You keep everything, including the battery.

Investment

One price. No scoping theatre.

Fixed scope, fixed price, one signature — approvable from a marketing, e-commerce, or comms budget without a procurement cycle. If it turns out models believe only true, flattering things about your brand, you'll have the evidence to show your board. That has not happened yet.

+ Additional market or brand — €2,900 each. Localized battery, same instrument, same report.

+ Canon & Repair Sprint — €12,000. We implement the Now-tier fixes with your team in three weeks: your machine-readable brand canon, product data harmonization, schema, and the highest-value source repairs. The same instrument we run on ourselves — see our own AI info page and capabilities.json.

+ Quarterly belief tracking — €3,500/quarter. The battery re-run: which false beliefs died, which narratives moved, your Belief Integrity Score and Share of Answer over time. Brand tracking for the model layer.

CEE pricing shown · EU/UK and enterprise pricing on request

€7,500
Fixed · one brand · one market
  • Share of Answer Scorecard, benchmarked
  • Model Belief Report with graded ledger & threat register
  • Claim-by-claim adjudication
  • Source Graph for every priority belief
  • Influence roadmap: now / next / sustained
  • 60-min executive readout · 10 working days
Who this is for

Built for brands whose claims do the selling

If your category is bought on trust in a claim — infant nutrition, supplements, functional foods and dairy, dermocosmetics, OTC, pet food, plant-based, bottled water — models are adjudicating those claims right now, and someone is winning the verdict. It's equally built for comparison-driven categories — appliances, baby gear, consumer electronics — where agent-built tables decide the purchase. Buyers: brand, e-commerce, and digital directors, plus the legal and comms leaders who discover mid-readout that this was their audit too.

Adjacent, but not the same

How this differs from what you may already run

Four adjacent things get confused with this audit. Each is useful; none answers the same question.

MethodPrimary questionEvidenceBoundary
SEO audit How do pages rank in traditional search? Technical, content and link factors Does not directly test model beliefs or recommendations
AI visibility / GEO tool Does the brand appear across prompts? Automated monitoring and share-of-answer metrics May not investigate claims, transcripts or source causality deeply
Brand tracking What do humans think? Survey and panel data Does not show what AI systems assert
Physical shelf audit Is the product available and correctly merchandised? Stores, facings, planograms, compliance A different shelf entirely
The Agentic Shelf Audit How do AI systems represent and judge the brand? Queries, transcripts, beliefs, claims, competitors and source tracing A focused, senior-led diagnostic and repair plan

↔ scroll table on mobile

The closest neighbour is the GEO monitoring tool, and the distinction is worth being precise about: a tool tells you your share of answer moved. This audit tells you what models believe, why they believe it, and which of those beliefs is worth the effort to change.

Ethics & boundaries

What we will not do

This audit is diagnostic. The repair work it recommends is about correcting the public record, not gaming it.

No astroturfing

We do not create or coordinate inauthentic third-party content.

No fake reviews or manufactured sources

Ever, in any form.

No prompt manipulation

We do not attempt to inject instructions into model behavior or exploit retrieval systems.

Substantiated information only

Corrections we recommend must be true and evidenced. If a model's unflattering claim is accurate, we will tell you that instead.

The durable position comes from being accurately represented, not from temporarily gaming a retrieval system that will change next quarter.

Common questions

Frequently asked

Which models do you cover?

The major assistants your buyers actually use. The exact set is agreed at market definition and recorded in the report, because coverage changes and comparability depends on knowing exactly what was tested.

Which markets and languages?

Models answer differently by language and region — often materially. Scope is agreed up front, and one audit covers one brand in one market.

Is it repeatable?

Yes, deliberately. You receive the full query battery, so the audit can be re-run and compared over time rather than being a one-off snapshot.

Can we run this ourselves?

Partly. You can ask a model about your brand today. What is harder is designing an unbiased battery, running it consistently across models, capturing findings in a form that survives scrutiny, and tracing assertions back to sources.

What if we disagree with a finding?

The belief ledger records what the model said, verbatim, with prompt and date. It is not our opinion of your brand — it is a record of what systems assert. Where a model is wrong, that is the finding.

Will you review our claims legally?

No. We flag claims that models contradict or hedge; your legal and regulatory review stays with you.

How long does it take, and what does it cost?

Ten working days, €7,500 fixed, one brand in one market. Fixed scope and fixed price, agreed before we start — no scoping theatre.

From the auxfirst canon

The thinking behind the instrument

Start here · Free

See three things AI models believe about your brand — before you spend a euro.

Send us your brand and market. Within five working days you'll get a three-belief teaser: real transcripts of what assistants currently assert, get wrong, or judge about your brand — each with the source it likely came from. No deck, no pitch, no obligation. If the findings don't unsettle you, delete the email.

Request your 3-belief teaser

In the message, mention "Shelf Audit teaser" plus your brand and market — that's all we need to start.

Capacity note Audits are booked one brand at a time to keep fieldwork honest and readouts senior-led. The teaser queue is capped at five brands per week. We audit and repair the information environment — we don't astroturf it. No fake reviews, no seeded forums, no prompt games: the lever is making the true, substantiated version of your brand the easiest thing for a model to find and cite.

Audits are led personally by Emil Krzemiński, founder of auxfirst, the agentic experience design agency — the practice that publishes its own machine-readable canon at auxfirst.com/ai-info.

The Agentic Shelf Audit™ · Model Belief Report™ · auxfirst service marks