Citeify

Methodology v2026.08

How the audit actually works

Most AI-visibility audits are one prompt to one model, dressed up as analysis. This page explains what ours measures, how it's scored, what we refuse to claim — and a dated changelog of every methodology change, so you can see the audit is maintained, not fossilised.

Maintained by James Eltherington, founder.

Measured AI visibility

We don't infer your AI visibility from on-page signals — we measure it. The audit builds a set of questions your customers actually ask, puts them to ChatGPT, Perplexity, Gemini and Google AI Overviews, and records the answers: whether you appear, where you rank, and who gets recommended instead of you. The queries and results are in your report — you can re-run any of them yourself and check.

Engines don't give the same answer twice, so a single query proves nothing — that's why this is a structured probe across engines and queries, scored in aggregate, rather than one screenshot of one lucky answer.

How scoring works

Your Readiness score is grounded in deterministic checks: measured signals from your live site, evaluated in code against a fixed rubric. The same site produces the same score, every run. On top of that, an AI analysis pass adds depth — but it is anchored to the measured signals, not free to improvise.

Your AI Visibility score comes from the measured probe above. The two blend into one Overall score across six categories:

We publish the categories, not the individual checks and weights — that's the part competitors would like us to hand over.

The QA gate & refund

Before a report reaches you it passes an automated QA gate: completeness of every section, presence and consistency of scores, and evidence backing the findings. A report that fails the gate is automatically retried; if it still can't pass, it is held for human review — and if we can't deliver, you get a full refund. You will never receive a report that didn't pass.

What we refuse to do

The GEO industry has a snake-oil problem. These are written rules enforced in our report pipeline — not marketing copy:

We run it on ourselves — current score: 41/100

Every methodology version is run against citeify.ai before it ships — same pipeline, same scoring, no special treatment. Our current result under v2026.08: Readiness 82/100, measured AI Visibility 0/100, Overall 41/100. Zero visibility means that across 120 recorded AI answers to buyer questions in our own niche, no engine cited citeify.ai. We publish that willingly: it is the honest starting point of a brand that is weeks old, and the number every future re-audit gets measured against.

The instrument itself is on display in the run history: two audits of the unchanged site scored an identical 76 Readiness — same site, same score — and after a set of fixes our own report recommended, the next run measured 82. Flat when nothing changes, moves when something does.

Read our full audit report (PDF) →

Methodology changelog

AI search changes monthly; an audit that doesn't is worthless. Every change to what we measure is versioned and dated here. Your PDF report carries the methodology version it was produced under.

v2026.08

August 2026

  • Unified the scoring model so every layer of the audit — analysis, report and PDF — is produced from a single documented rubric.
  • Visibility probe now samples each question multiple times per engine (AI answers vary run to run — one sample is a coin flip) and re-uses a domain’s question set between audits, so score changes measure visibility movement, not question drift.
  • Tightened our evidence policy: every claim in a report must trace to a measured signal or published research. Estimated traffic and revenue projections are now prohibited in our reports.
  • Recalibrated research citations against the current academic literature on generative engine optimisation (2024–2026).
  • Every report PDF now carries its methodology version, tied to this changelog.

v2026.07

July 2026

  • Added the measured AI-visibility probe: your customers’ real questions are put to ChatGPT, Perplexity, Gemini and Google AI Overviews, and the answers recorded.
  • Moved readiness scoring to deterministic, code-computed checks grounded in measured site signals — the same site now always produces the same score.
  • Added multi-page citability analysis and automated QA gating on every report.

See it applied to your site.

Around 15 minutes from domain to report. QA-gated, refund-backed.

Audit Methodology — Citeify