Worked example°

A full AI Visibility Audit, redacted

This is what the free audit produces: the prompt set, the observed responses, the competitor set, the scoring against a published standard, and the roadmap. Published so the “do you actually know how to do this” question is settled before the first call rather than during it.
A team reviewing the findings of an AI visibility audit
Subject
A mid-market API observability platform
Window
1 to 14 July 2026
Runs per prompt
5

What is real here, and what is not

The subject is an anonymised composite of engagements in one category, not a single named client. The competitor names are redacted. The failure modes, the source concentration and the shape of the roadmap are what these audits actually surface. The individual figures should be read as illustrative of the method, not as a verifiable result for a specific company. We would rather say that plainly than let you assume otherwise.

Section one

The prompts, and what came back

Six of the forty-two prompts tested, chosen because each shows a different failure. Note the follow-up turn: single-prompt testing would have missed the most important finding in the audit.

Category discoveryAbsent

What are the best API observability tools for a mid-sized engineering team?

Five products named. The subject was not among them. Three of the four cited sources were third-party; the only vendor domain cited belonged to a competitor whose documentation answered the question directly.

Sources cited:1. reddit.com2. g2.com3. techradar.com4. datadoghq.com
Category discoveryMentioned

Which observability tools work well with OpenTelemetry without vendor lock-in?

Subject named fourth of five, described accurately but briefly. No link. The description came from a Reddit thread rather than from the subject's own site, which had a stronger page on the same topic that was not retrieved.

Sources cited:1. reddit.com2. opentelemetry.io3. grafana.com
ComparisonCited

[Subject] vs [Competitor A]: which is better for a Series B startup?

Balanced comparison, subject's own pricing page cited. One material error: a pricing tier that was discontinued nine months earlier was presented as current, sourced from a G2 listing that had never been updated.

Sources cited:1. g2.com2. [subject].com3. [competitor-a].com4. reddit.com
QualificationCited

What does [Subject] cost?

Correct on the current published tiers, then repeated the discontinued tier from the same stale G2 listing as a third option. Brand accuracy failure, caused by a third-party source the subject controls the content of.

Sources cited:1. [subject].com2. g2.com
Follow-upAbsent

Follow-up: which of those would you pick for a team of 30 engineers?

Subject dropped out entirely at the narrowing turn. The retrieved set moved to team-size-specific content, which the subject had none of. This is the most common failure we see and it is invisible to single-prompt testing.

Sources cited:1. reddit.com2. [competitor-b].com3. capterra.com
QualificationAbsent

Which API observability tools have the best onboarding for small teams?

A YouTube walkthrough by an independent practitioner was cited second. The subject had no third-party video coverage at all, on a question their product genuinely wins.

Sources cited:1. reddit.com2. youtube.com3. [competitor-b].com
Section two

The scoring

Every metric below is defined in the published measurement standard, with its formula and what it excludes. Citation rate and brand mention rate are reported separately, never blended.

42 prompts
Prompt coverage

Declared and frozen before measurement: 12 category discovery, 14 comparison, 10 qualification, 6 follow-up turns. Each run 5 times across the window.

19%
Citation rate

40 of 210 answers linked a page on the subject's domain. Concentrated almost entirely in branded comparison and pricing prompts.

34%
Brand mention rate

71 of 210 answers named the subject. Reported separately from citation rate, never blended, because blending them would nearly double the headline number.

11%
Share of voice

Against a competitor set of six, declared before measurement. The category leader took 31%.

72%
Brand accuracy

20 of 71 answers naming the subject contained a material error. 17 of those 20 traced to one stale G2 listing.

Reddit 22%, G2 17%
Source concentration

Of all citations across the set, the subject's own domain accounted for 9%. The rest sat on sources the subject did not own and was not working on.

Citation rate by platform, never averaged
PlatformCitation rateWhat explains it
ChatGPT16%Heaviest reliance on Reddit of any platform tested. Subject's absence from community threads cost most here.
Gemini24%Best performance, tracking the subject's existing Google rankings. Mention rate far exceeded citation rate, as expected on this surface.
Claude27%Strong where documentation answered the question. The subject's docs are public and good, which is why this is the best citation rate in the set.
Perplexity14%Visible source lists made the concentration on Reddit and G2 directly observable, which is what directed the roadmap below.
Copilot9%Traced to weak Bing indexation rather than anything AI-specific. 40% of the subject's pages were not in the Bing index at all.
Section three

The roadmap that comes out of it

Ordered by damage done, not by effort. Note that the top two items are days and weeks of unglamorous work, and neither is content production.

1Days

Correct the stale G2 listing

One out-of-date third-party listing caused 17 of the 20 material accuracy failures. The subject controls this content and had not looked at it in two years.

2Weeks

Fix Bing indexation

40% of pages missing from the index Copilot grounds on. Unglamorous, cheap, and the entire Copilot problem.

3Weeks

Publish team-size and stage-specific qualification content

Complete drop-out at the narrowing turn. The retrieved set moves to content the subject does not have, on questions their product wins.

4Quarters

Engage the twelve Reddit threads that carry category citations

Reddit accounted for 22% of all citations. Transparent factual correction first, then genuine practitioner participation.

5Quarters

Support independent video coverage of onboarding

A competitor's YouTube walkthrough is cited on a question the subject wins on the merits. There is no subject-related video in the retrieved set at all.

Illustrative composite. These are engagement patterns described by category, not named clients, and the figures show the shape of an outcome rather than a single verifiable result. Read them as an indication of what the work produces, not as a claim you could check. Where we publish a number about the firm itself, it carries its sample size, window and method on the page.

The same thing, for your brand

The free AI Visibility Audit produces this, on your category, against your competitor set, scored against the same published standard.