Buyer questions
A versioned set of real category, comparison, problem and buying questions. Branded checks are reported separately so they do not inflate discovery visibility.
AI Search measurement
AI Search measurement is a repeated sample of defined buyer questions across named engines, markets, and dates. We keep brand mentions, competitor share, answer position, citations, retrieval, sentiment, factual accuracy and commercial outcomes separate. A change in one metric does not stand in for another, and movement after our work does not by itself prove causation.

The short version
A credible report shows the question set, engines, country, language, sample dates and valid-answer count. It defines every metric before showing a score. It also marks missing runs and method changes, because a neat percentage built on changing coverage is not a clean comparison.
We use scheduled observations to catch changes, then compare completed periods defined in the written measurement plan. The buyer can see what moved, what evidence supports the interpretation, what remains unavailable and which next action is justified.
01
The measurement frame is agreed before the baseline. Each response remains an observation tied to its settings, not a timeless ranking.

A versioned set of real category, comparison, problem and buying questions. Branded checks are reported separately so they do not inflate discovery visibility.
Only the named engines and answer surfaces in scope. Results are broken out by engine before any combined view.
Country, language and locale are fixed where the engine allows it. A UK English result is not silently mixed with a German result.
Before the baseline, the plan states the exact repeat count for each question, engine, market and language cell, the schedule, the period length and the minimum valid coverage needed for comparison. Every run records its date and outcome.
We retain the answer, brand and competitor mentions, visible order, linked or cited sources, and the factual statements needed for accuracy review.
A valid answer is a completed in-scope response that can be stored and reviewed. The plan freezes cell weights, inclusion rules, metric formulas, evidence authority, low-volume threshold and any uncertainty method before the baseline.
02
Each layer answers a different question. The report keeps the numerator, denominator and unavailable observations visible so the score can be checked.
| Layer | Definition | What it can tell you | What it cannot prove |
|---|---|---|---|
| Visibility | Valid answers naming the brand at least once divided by all valid answers in the frozen cell set. One answer contributes at most one brand-presence count. | Whether the brand appears for the tracked question set. | Prominence, preference, a citation or a visit. |
| Share of voice | Brand-presence credits divided by all brand and frozen-competitor presence credits. Each tracked brand contributes at most one credit per answer; the report shows both counts. | Relative presence inside the defined competitor set. | Market share, buyer preference or revenue share. |
| Position | For valid answers that both name the brand and contain a meaningful ordered list, record the first visible brand position. Report the eligible-answer count, position distribution and median, not a rank for ineligible answers. | Whether the brand tends to appear early or late when order is observable. | A universal rank. Many answers have no meaningful ordered list. |
| Citations and retrieval | A citation is a visible supporting link. Retrieval is a separately recorded source-use event only when the measurement system exposes it, whether or not that source is also cited. Retrieved-but-not-cited is a reported subset. | Which domains and pages support or are exposed as inputs to sampled answers, and where source gaps exist. | That source use caused the brand mention, or that an unexposed retrieval event occurred. |
| Sentiment | Each brand-bearing answer is labeled positive, neutral, negative or mixed under the versioned rubric. The report shows category counts and shares, retains the passage and sends ambiguous cases for review. | Whether recurring themes appear in the sample. | Customer satisfaction or reputation across the whole market. |
| Accuracy | Correct checkable claim occurrences divided by all checkable claim occurrences about the brand. The plan names the approved source of truth; disputed and uncheckable claims stay outside the denominator. | Where engines repeat outdated, incomplete or incorrect information. | That silence is accurate, or that a correct claim persuaded the buyer. |
| Qualified outcomes | A funnel, not one blended score. Report separate counts and conversion rates for identifiable engaged visits, inquiries, booked calls and opportunities that pass the pre-agreed qualification rule, with unknown source retained as unknown. | Whether measurable commercial activity followed the exposure path. | That AI visibility alone caused the outcome. |
03
AI answers can change when models, source indexes, location, personalization or the question wording change. We therefore preserve the measurement frame and report the actual coverage beside every comparison.

Version the questions, competitors, engines, markets, cell weights, repeats, period length, formulas and decision threshold before the baseline.
Count valid answers and label unavailable, failed or unsupported observations. Missing data is not zero.
Give each frozen cell its predeclared weight. Use 95% Wilson intervals only for unweighted answer-level proportions that meet the plan's independence rule. Correlated repeats or weighted estimates need a separately frozen cluster-aware or design-based variance method; otherwise label the interval unavailable and the estimate descriptive.
Inspect engine, topic, funnel stage and market separately. A combined score can hide one strong engine and one failing engine.
Record when a page, source, prompt set, tracking rule or model changes. The timeline makes interpretation possible without pretending it isolates cause.
04
The strongest claim must match the strongest evidence available.
A metric moved after a dated change. This is a signal worth investigating.
A visit or outcome carries a source, campaign or declared discovery path under the chosen analytics rules. Credit depends on the attribution model and lookback window.
Answer visibility, referred visits and qualified outcomes move in a consistent direction across repeated periods, with major confounders reviewed.
The evidence isolates the intervention from other plausible causes, usually through a suitable experiment or control. A before-and-after chart alone does not meet this bar.
05
We join the journey only where the evidence allows it. Answer data shows exposure. Analytics shows identifiable visits and on-site actions. Forms and booking systems show inquiries and appointments. Qualification belongs in the CRM or an agreed review, not in a pageview count.

Brand visibility, share of voice, position and supporting sources.
Identifiable referred visits, methodology-page engagement and audit transitions.
Submitted audits, contact forms or booked-call starts with safe campaign fields.
Accepted fit criteria, opportunity status and disqualification reason where available.
Won work or another agreed commercial result, reported separately from attributed influence.
07
These sources explain platform mechanics. They do not endorse Schmitdy or guarantee visibility.

FAQ
No. Visibility measures whether a brand appears in sampled answers. Traffic measures identifiable visits. An answer can influence a buyer without a click, while a visit can arrive without a measured brand mention.
Not cleanly. Share of voice depends on the questions, engines, competitor set, market and valid-answer coverage. If any of those change, the report must label the new version and avoid presenting the movement as a like-for-like trend.
Not necessarily. Position records visible order only when order is meaningful. Recommendation strength, wording, citations and accuracy need separate review.
No. A citation is a visible supporting link. Retrieval is an internal use of a source and is recorded only when the measurement source exposes it. We never infer hidden retrieval from a mention alone.
The report retains the relevant answer passage, assigns a defined label and reviews recurring themes by topic. The score is an interpretation of the sample, not a customer-satisfaction survey.
A checkable statement must match current approved evidence. Outdated, incomplete and incorrect statements are separated, and disputed or uncheckable claims stay outside the accuracy denominator.
Only to the level the evidence supports. Tagged visits, declared discovery paths and CRM qualification can support attribution. A visibility increase followed by revenue is still not causal proof without a design that rules out other plausible causes.
The first report should show the frozen measurement frame, baseline counts, metric definitions, missing-data rules, source examples, qualification rules and the next observation window. If the current baseline is unavailable, it should say unavailable rather than record zero.
Your existing stack
The method works with the website, analytics and CRM stack you already use. We agree the evidence path before claiming an outcome.





Next step
The free audit maps your buyer questions, current answer coverage and the clearest source gaps. No invented baseline and no promise that one edit caused the result.
Get your free AI Search audit β