Playbooks · Playbook

GEO Audit Playbook

A 1–2 day GEO audit playbook: prepare scope, capture AI answers, tag citations, prioritize gaps, and remeasure.

Updated Sep 7, 2026 Reviewed Sep 7, 2026 en

Run this playbook when leadership asks whether the brand is visible in AI answers but the team does not yet have repeatable evidence.

It is the operator version of the GEO audit workflow. The guide explains the sequence. This playbook tells a small team how to finish a first audit in one to two working days without turning the work into an open-ended research project.

Keep the scope narrow. One category, one market, one buyer role, and a short competitor list will produce a usable decision. A 40-prompt, five-market sweep will not, not on this timeline.

What this audit is for

A GEO audit answers four questions with evidence, not opinions:

  1. Which buyer questions currently produce answers about this category?
  2. Does the brand appear, and how is it framed?
  3. Which sources are cited, and do owned pages appear among them?
  4. Which competitors occupy the recommendation, comparison, or “safe default” slots?

Use Prompt, Presence, Proof, Position as the tagging language. Do not wait for a perfect measurement stack. A shared sheet plus stable prompts is enough for the first pass.

Day 0 to morning of day 1: prep

Do not start prompting until the inventory is written down. Prep usually takes 60–90 minutes.

Prep itemDone when
Category and buyerYou can name one category, one market, and one primary buyer task.
Prompt themesYou have at least definition, category, comparison, recommendation, and problem-solution themes.
Competitor shortlistYou listed 3–5 names the buyer would actually compare, not every adjacent brand.
Owned sourcesYou listed the URLs that should be citable: category page, comparison page, docs, glossary, proof page.
SurfacesYou named the AI answer surfaces the team actually cares about this week.
Capture sheetColumns exist for prompt, surface, date, answer summary, brand, competitors, citations, accuracy, action.

If the owned-source list is empty, stop and collect it. An audit that cannot map an answer back to a page cannot produce a fix list.

Prompt set

Build 12–20 prompts for the first pass. Fewer than 10 leaves too many holes. More than 25 usually delays tagging.

Cover these jobs:

Prompt jobExample shapeWhy it belongs
Definition“What is [category]?”Shows whether the category is explained in language the team would accept.
Category discovery“Best tools for [job]”Shows who occupies the shortlist.
Comparison“[Brand] vs [competitor]”Shows framing, not only presence.
Recommendation“Which [category] should a [role] use?”Shows whether the brand is treated as a default, alternative, or omission.
Problem-solution“How do I [buyer task]?”Shows whether owned how-to sources can enter the answer.
Accuracy trap“Does [brand] support [claim]?”Shows stale or invented product descriptions.

Write the exact wording in the sheet and freeze it for this audit. Small wording changes are a later experiment, not part of the first capture. See prompt tracking for why the string itself is evidence.

Capture

Run the same prompt set across the named surfaces. Capture on the same day if possible so the batch is comparable.

For every row, save:

Do not score yet. Capture first. Scoring during capture usually collapses the record into a vibe.

If a surface shows no citations, record that as a finding. Missing attribution is still evidence; it just changes how much weight the row can carry for source work.

Tag mentions, citations, and competitors

Tagging is the actual analysis. Budget most of day 1 afternoon for it.

Use a consistent code, then write one sentence of context:

TagMeaning
Present / omitted / misdescribedBrand presence in the answer body.
Cited owned / cited third-party / no citationWhether a source is visible, and whose it is.
Lead / co-mentioned / alternative / absentCompetitor or brand position in the recommendation frame.
Accurate / stale / inventedWhether the answer matches current public facts the team can defend.

Then connect each weak row to a page or proof gap:

Use the content citability checklist when the issue is extractability rather than missing coverage.

Prioritize fixes

Do not produce a 30-item content backlog. Rank a short action list that a team can start this week.

Score each issue on:

  1. Prompt importance: is this a buyer question the team already pays to win in search?
  2. Frequency: does the gap repeat across prompts or surfaces?
  3. Fixability: is there an owned page, proof asset, or entity-consistency issue the team can change?
  4. Risk: is the answer actively wrong about product, pricing, or category?

A practical priority order:

PriorityTypical patternFirst fix
P1High-intent recommendation or comparison prompts omit or misdescribe the brandRepair the owned category, comparison, or product-fact page
P2Brand appears, but citations go to competitor or generic sourcesStrengthen extractable claims, dates, and source clarity on owned URLs
P3Definition prompts describe the category looselyPublish or tighten a durable explainer that matches buyer language
P4One-off odd answers with no repeat patternLog for the next measurement pass; do not rebuild the site around them

The output of this step is the expected measurement outcome: a ranked list of prompts, citation gaps, competitor visibility patterns, and source improvement tasks.

Re-measure

Re-measure only after a defined change, not after “we talked about GEO.”

Wait until at least one P1 or P2 page has been updated, then rerun the same prompt set on the same surfaces. Keep the original rows. Add a second capture date. Compare presence, citations, framing, and accuracy notes.

If nothing changed, that is still a result. It usually means the wrong page was edited, the claim is not corroborated off-site, or the prompt set was too brand-centric. Fold that lesson into how to measure AI visibility rather than widening the audit immediately.

Output format

Create one audit table with these columns:

ColumnUse
PromptFrozen wording from the set
SurfaceNamed AI answer surface
Answer summaryThe claim that matters, not the full transcript
Brand presencePresent, omitted, or misdescribed
Competitor presenceWho else appears, and in what frame
CitationsOwned, third-party, or none
Accuracy notesStale, invented, or acceptable
Recommended actionOne page, proof, or entity fix, with a priority

Attach a one-page memo: scope, date range, top five findings, and the remeasure date. That package is enough for an SEO operator, a growth lead, or an agency client review.

Operator checklist

If the team later needs a longer validation cycle rather than a first snapshot, move to the 90-day operating rhythm. If the question is which measurement jobs tools should support, use the 2026 tool landscape.