Visibility
What AI engines see when asked about your brand
Sourced from live probes across 5 AI engines. AI answers vary by day and engine.
Fraction of buyer questions where an AI engine mentioned your brand. We ask each buyer question in 2 different wordings, and we run each wording 2 times per engine. That is 4 runs per question, per engine, before we add anything. From those runs we compute a mention rate, not a single coin flip.
When cited, how high in the AI answer? Position 1 = 1.0, position 2 = 0.5, position 3 = 0.33, and so on. Zero if never cited.
We classify the text around each brand mention as positive, neutral, or negative using a deterministic phrase-matching classifier. Positive = 1.0, neutral = 0.5, negative = 0.
Why we run each question more than once
AI engines are not consistent. The same question can get a different answer on the next request. So we never judge your brand on one run.
We start lean. Each buyer question is asked in 2 wordings, and each wording runs 2 times per engine. That is a base of 4 runs per question, per engine.
Then we only spend more where the answer is unclear. If a question lands in the grey zone on an engine, meaning your brand is cited in 25% to 75% of the runs, we add 1 more run per wording. We do that for at most 2 extra rounds, and we stop once that question reaches 6 runs on that engine. A clean 0 of 4 or 4 of 4 stays at 4 runs. Paying for more runs would not tell you anything new.
Every audit has a ceiling. There is a hard limit on how many AI runs one audit can spend (220 by default). If an extra round would cross it, we stop adding runs instead of cutting corners elsewhere. The base 4 runs always finish.
Every rate ships with its margin of error. We report a 95% Wilson confidence interval next to each rate, so you can see how solid it is. 4 of 4 runs gives a tight range. 2 of 4 gives a wide one, and that width is the honest part. When only 1 run exists (older audits), we show “single sample” instead of claiming confidence we do not have.
Honest note: AI outputs are not deterministic. The same prompt can produce different answers on different days. We run each buyer question at least 4 times per engine, add runs where the signal is unclear, and publish a confidence interval with every rate. It is still a snapshot. Re-audit weekly to track the trend, rather than treating any single score as absolute.