GEO checklist

Every GEO readiness check Seov runs.

GEO measures technical readiness for AI crawlers and answer engines — it does not measure or predict whether any specific AI system currently cites your site. These are the real 15 checks.

AI crawler access (9)

Major AI crawlers not blockedVery important
Rolled-up summary of the per-bot robots.txt breakdown below — GOOD if no major AI crawler is blocked, WARN if some are, CRITICAL if all are (or robots.txt disallows "*" outright).
Fix: Remove blanket Disallow rules for AI crawlers (GPTBot, ClaudeBot, PerplexityBot, etc.) on public content.
GPTBot (OpenAI)Low importance
Informational: whether robots.txt allows OpenAI's GPTBot.
ChatGPT-User (OpenAI)Low importance
Informational: whether robots.txt allows OpenAI's ChatGPT-User (live browsing on a user's behalf).
Google-Extended (Gemini/Bard training)Low importance
Informational: whether robots.txt allows Google-Extended, distinct from Googlebot search indexing.
ClaudeBot (Anthropic)Low importance
Informational: whether robots.txt allows Anthropic's ClaudeBot.
anthropic-ai (Anthropic)Low importance
Informational: whether robots.txt allows Anthropic's anthropic-ai user agent.
PerplexityBot (Perplexity)Low importance
Informational: whether robots.txt allows Perplexity's crawler.
CCBot (Common Crawl)Low importance
Informational: whether robots.txt allows Common Crawl's bot — a dataset several AI labs train on.
Applebot-Extended (Apple)Low importance
Informational: whether robots.txt allows Applebot-Extended, distinct from Applebot search indexing.

Structured data (1)

JSON-LD structured data presentImportant
Schema.org JSON-LD (Organization, Article, FAQPage, Product, etc.) helps machines understand page content unambiguously.
Fix: Add Organization schema sitewide and content-specific schema (FAQPage, Article, Product) where relevant.

Extractability (3)

llms.txt presentImportant
An emerging convention giving AI systems a machine-readable summary of the site. One signal among several, not a guaranteed ranking/citation factor.
Fix: Add a /llms.txt describing the company, key pages, and citation policy.
Meaningful content present in server-rendered HTMLImportant
AI crawlers generally don't execute JavaScript — if the initial HTML response is mostly empty markup with content injected client-side, there is little for them to extract regardless of how well the site reads in a browser.
Fix: Server-render (or statically generate) primary content instead of relying on client-side JavaScript to populate it.
Clear semantic HTML structureLow importance
Use of <main>/<article> and a reasonable text-to-markup ratio makes content easier for machines to extract.
Fix: Wrap primary content in semantic elements and reduce unnecessary markup nesting.

Entity clarity (2)

Organization entity clearly definedImportant
Organization schema and consistent naming help AI systems correctly identify who the site belongs to. Heuristic/directional signal.
Fix: Add Organization JSON-LD with a consistent name, logo, and social profile links.
Page language declaredLow importance
Informational: the language declared in the HTML lang attribute.