Guides · Guide

GEO Audit Workflow

An operator workflow for scoping, capturing, tagging, and acting on GEO audits of brand visibility, citations, and answer quality.

Updated Sep 7, 2026 Reviewed Sep 7, 2026 en

A GEO audit checks how AI answer systems describe a brand, category, and competitors today. The useful output is evidence: what the answers say, which sources they cite, where the brand is missing or misrepresented, and which pages can change that.

Use the GEO audit playbook for a first-run table. Use how to measure AI visibility when the audit should become a recurring program.

1. Scope the audit

Do not start with every prompt and every surface. Record these decisions before capture:

Scope itemDecision to record
CategoryThe market or problem space the buyer is researching.
Brand and product namesExact names, former names, and names to ignore.
Competitor setA shortlist, not every adjacent vendor.
Markets and languagesCountry, language, and whether localized answers are in scope.
SurfacesWhich AI search, answer engine, or assistant experiences will be captured.
TimeboxHow many prompts, how many surfaces, and when capture closes.

A focused first audit often covers one category, 20 to 40 prompts, two or three surfaces, and five to eight competitors. Broader programs come after the first evidence set is usable.

Counterexample: collecting twenty random chats from different people, then reporting “we are not visible.” That is a conversation. The prompt, surface, date, and competitor set were never controlled.

2. Build the prompt set

The prompt set is the measurement object. Prompt tracking only works if the questions reflect real discovery and evaluation work.

Cover these prompt types:

TypeWhat it testsExample shape
DefinitionCategory language and whether owned sources explain the topic.”What is generative engine optimization?”
CategoryHow the market is framed and which tool types appear.”What are GEO tools?”
RecommendationBrand presence, order, and shortlist quality.”Which AI visibility tools should a B2B SaaS team consider?”
ComparisonHow alternatives are described together.”GEO tools vs SEO rank trackers”
WorkflowWhether owned guides appear as practical sources.”How should an agency run a GEO audit?”
Problem-solutionWhether the brand is attached to the job, not only the category name.”How do I check if an AI answer cites my site?”

Keep core prompts stable. Put new buyer phrasing in a backlog so the trend is not destroyed by constant rewording. SEO demand is an input, not a keyword paste: map impression-heavy queries into the same intent families, then add conversational variants.

3. Capture protocol

Capture is a protocol, not a screenshot dump:

  1. Use the recorded prompt wording. Do not paraphrase mid-run.
  2. One prompt per record. Do not merge follow-up chat unless that follow-up is a tracked prompt.
  3. Record surface, logged-in or logged-out state, market, and language.
  4. Save the full answer text and visible citations exactly as shown.
  5. Record the capture date. Do not edit the answer before tagging.

If a surface shows no citations, still keep the answer. Uncited answers are evidence. They often reveal framing problems that citation counts miss. See citation tracking.

4. Tagging taxonomy

Tagging turns answer text into comparable fields. Use a small, consistent taxonomy.

TagValuesHow to use it
Brand presencementioned / omitted / unclearUnclear is for ambiguous nicknames or product-line confusion.
Mention rolerecommended / compared / defined / passing / negativeA passing mention is not a win.
Positionfirst / middle / last / not applicableUse only when the answer ranks options.
Accuracyaccurate / stale / incomplete / incorrect / mixedMixed is common; do not force a single score.
Citation ownerowned / competitor / third-party / uncitedTrack the supporting source, not only the brand name.
Competitor hitwhich shortlist names appearRelative visibility is the audit, not an isolated mention.

Do not invent extra tags during the first audit. If a new distinction keeps appearing, add it next cycle and re-tag the baseline.

Example: the brand is third in a five-name shortlist, described with last year’s category, and cited only via a partner blog. Presence is true. The action is to fix category language and owned-source citability, not to celebrate the mention.

5. Evidence fields

Keep enough fields that a later run can be compared without tribal knowledge.

FieldWhy it is required
Prompt ID and wordingTies the record to the controlled set.
Intent familyGroups definition, category, comparison, and workflow prompts.
Surface, market, language, dateMakes runs comparable across systems and locales.
Full answer textScores cannot reconstruct claims.
Brand mention and roleSeparates presence from recommendation quality.
Competitor names and orderExplains relative visibility.
Cited URLs and source ownerShows who is supporting the answer.
Accuracy notesRecords stale, missing, or unfair claims.
Follow-up actionConnects evidence to a page, source, or entity fix.

If a field cannot be filled, write unknown rather than guessing. Unknown capture conditions are themselves an audit finding.

6. Output an action list

The audit is finished when it produces an action list, not when it produces a score.

Group actions by the thing an editor or SEO can change:

FindingTypical action
Brand omitted on category and recommendation promptsClarify category pages, comparison assets, and entity language.
Brand mentioned, owned pages never citedImprove citability: definitions, evidence, headings, and source freshness.
Third-party pages define the categoryStrengthen owned glossary and guide pages that answer the same question.
Competitors recommended first with better comparison pagesBuild or refresh a comparison that names alternatives fairly.
Answer is accurate but staleUpdate product, category, and positioning language across owned sources.
Answer is incorrectFix the owned source of truth, then check which third-party pages still repeat the error.

Each action should name an owner, a URL or source type, and the prompt family it is meant to change. “Create more content” is not an action.

7. Cadence

A first audit is a baseline. Re-run the same core prompt set after the first action batch lands.

SituationPractical cadence
First baselineOne complete capture window, then a re-run after the first fixes.
Stable category, small teamMonthly review of the core set.
Competitive category or frequent product changesWeekly or biweekly on recommendation and comparison prompts.
After a launch, repositioning, or major content refreshImmediate targeted re-run, then return to the regular cadence.

Changing the prompt set every cycle destroys the trend. Add prompts in a named backlog and promote them into the core set only when they will be kept.

Common audit failures

Next step

Re-run from the GEO audit playbook table, then keep the loop in how to measure AI visibility and citation tracking.