AEO AND GEO SKILL
seo-geo
A scored GEO analysis of one URL: citability, structure, multimodal content, authority, and technical access, with the most accurate AI-crawler table we found on any shelf.
by Daniel AgriciAgriciDaniel/claude-seoMIT licencev2.3.1upstream 2026-09-10read by us 11 September 2026
It opens with Google’s AI optimisation guide as the primary source and frames every finding as SEO applied to AI surfaces, not a new discipline. When community advice contradicts Google, the skill says to defer to Google and note the contradiction in the report.
Five weighted criteria produce a GEO readiness score: citability 25%, structure 20%, multimodal 15%, authority and brand 20%, technical accessibility 20%. Each has strong and weak signals written out.
The crawler section separates training bots from search-citability bots for OpenAI, Anthropic, Google and Apple, with a table of which claim each user agent can and cannot support. This is the section to read even if you never run the skill.
When to use it
The phrases that trigger it
From the skill’s own description: say any of these and an agent that has it installed will load it.
- AI Overviews
- SGE
- GEO
- AI search
- LLM optimization
- Perplexity
- AI citations
- ChatGPT search
- AI visibility
What’s inside
The playbook, section by section
Key statistics with their sources named
Reach and coverage figures for AI Overviews and AI Mode, each tagged as Google-confirmed or third-party reporting.
Brand mentions over backlinks
The Ahrefs 75,000-brand study: YouTube mentions correlate at about 0.74 with AI citations, domain rating at about 0.27.
The crawler claim table
Eight claims, each paired with the one bot that supports it and the one that does not.
llms.txt evidence
A reference file collecting Mueller, Illyes, the SE Ranking 300k-domain study and a server-log audit: presence is reported, no citation weight is assigned.
Two Google citation engines
AI Overviews and AI Mode agree on the conclusion about 86% of the time and cite the same URLs 13.7% of the time (Ahrefs, 540k query pairs), so both are scored.
Quick wins, medium effort, high impact
Three tiers of changes, from a definition in the first 60 words to original research and Wikipedia presence.
Reference files: google-ai-optimization-guide.md llmstxt-evidence.md
See it in action
One real run, cropped
We installed the skill folder, gave it the request below, and kept what it returned. Nothing was edited; the excerpt is the first 1,583 of 15,474 characters. The full output is in our repository as research/skills/runs/seo-geo.md.
Run the GEO analysis on https://addition-labs.com/resources/guides/generative-engine-optimization. No DataForSEO. Produce the GEO-ANALYSIS output.
GEO-ANALYSIS
URL: https://addition-labs.com/resources/guides/generative-engine-optimization Fetched: 2026-09-11 (raw HTTP + Playwright render via claude-seo runtime; robots.txt, llms.txt, llms-full.txt, sitemap.xml also fetched) Framing: per Google's AI optimization guide, this is SEO applied to AI-search surfaces, not a separate discipline. Scores below are claude-seo heuristics, not Google-internal signals (Google, "Using third-party SEO tools", 2026-06-05).
1. GEO Readiness Score: 77/100
| Criterion | Weight | Score | Basis |
|---|---|---|---|
| Citability | 25 | 21 | Definition in first 48 words; 16 attributed primary sources; unique data (vendor unit comparison, Cloudflare Radar readings); three blocks in the 134-167 word band |
| Structural readability | 20 | 16 | Clean H1 > 9 H2 > 6 H3, TOC, 52 short paragraphs (median ~47 words), 1 table, 3 ordered lists, HTML FAQ. Only the FAQ headings are questions |
| Multi-modal | 15 | 9 | 8 figures with descriptive alt, 1 data table. No video, no embedded interactive element. 7 of 8 images lack width/height |
| Authority and brand signals | 20 | 13 | Byline is organizational ("By the Addition team"), no named person. Updated date visible (10 Sep 2026); publish date (25 Jul 2026) only in schema. Wikipedia/Wikidata: absent. LinkedIn: present. Reddit/YouTube: not measurable in this run |
| Technical accessibility | 20 | 18 | Full server-side rendering; robots.txt allows every crawler; llms.txt and llms-full.txt present and list this page; no RSL licensing |
2. Platform Breakdown
Cropped here. The rest continues in the same register.
What it could not do in this run
ranking/SERP data (no DataForSEO per request, no Search Console), Reddit mentions (403), YouTube mentions (script needs GOOGLE_API_KEY), IndexNow/Bing status, CWV field data; reported as not measured rather than estimated.
Method behind it
Where we would differ, and why
Our GEO guide reproduces the citation-overlap and recency findings this skill relies on, with the sample sizes beside each. Where the skill quotes a number, the guide shows where it came from and what it does not cover.
- Note 1
- The 134-to-167-word passage length and the 156% multimodal figure are quoted without a source in the skill text. We could not trace either to a primary study; treat them as the author’s rules of thumb.
- Note 2
- The skill still recommends creating llms.txt as a medium-effort item after its own reference file concludes it carries no citation weight. Do it last, if at all.
SKILL.md
The upstream file, as we read it
Copyright Daniel Agrici, MIT licence, commit 55c7914 of AgriciDaniel/claude-seo. Reproduced here under that licence so you can read what the agent will read; the folder’s reference files are in the repository.
Open SKILL.md (2,719 words)
---
name: seo-geo
description: >
Optimize content for AI Overviews (formerly SGE), ChatGPT web search,
Perplexity, and other AI-powered search experiences. Generative Engine
Optimization (GEO) analysis including brand mention signals, AI crawler
accessibility, llms.txt compliance, passage-level citability scoring, and
platform-specific optimization. Use when user says "AI Overviews", "SGE",
"GEO", "AI search", "LLM optimization", "Perplexity", "AI citations",
"ChatGPT search", or "AI visibility".
user-invocable: true
argument-hint: "[url]"
license: MIT
metadata:
author: AgriciDaniel
version: "2.3.1"
category: seo
---
# AI Search / GEO Optimization (May 2026)
## Primary Source: Google's AI Optimization Guide
Google's official position, published under Search Central docs:
> "Optimizing for generative AI search is **still SEO** from Google's
> perspective. AEO and GEO are rebranded labels for the same work."
Read `references/google-ai-optimization-guide.md` for the full synthesis,
myth-busting list (`llms.txt`, chunking, AI-rephrasing, mention-farming,
all rejected by Google as ineffective), and the Who/How/Why test for
content quality.
Audits should frame GEO findings as **SEO fundamentals applied to AI-search
surfaces**, not as a separate optimization discipline. When community
recommendations contradict Google's primary source, defer to Google and note
the contradiction in the report.
## Key Statistics
| Metric | Value | Source |
|--------|-------|--------|
| AI Overviews reach | 2.5 billion+ monthly active users, reported from Google I/O 2026 keynote coverage; not confirmed on a Google-owned source; 200+ countries | Third-party I/O reporting |
| AI Overviews query coverage | ~50% of queries (third-party measurement; varies by country) | Industry data |
| AI Mode monthly users | 1B+, reported from Google I/O 2026 keynote coverage; not confirmed on a Google-owned source | Third-party I/O reporting |
| AI Mode model | custom version of Gemini 2.5 | Google |
| AI-referred sessions growth | 527% (Jan-May 2025) | SparkToro |
| ChatGPT weekly active users | 900 million | OpenAI |
| Perplexity monthly queries | 500+ million | Perplexity |
## Critical Insight: Brand Mentions > Backlinks
**Brand mentions correlate 3x more strongly with AI visibility than backlinks.**
(Ahrefs December 2025 study of 75,000 brands)
| Signal | Correlation with AI Citations |
|--------|------------------------------|
| YouTube mentions | ~0.737 (strongest) |
| Reddit mentions | High |
| Wikipedia presence | High |
| LinkedIn presence | Moderate |
| Domain Rating (backlinks) | ~0.266 (weak) |
**Only 11% of domains** are cited by both ChatGPT and Google AI Overviews for the same query, so platform-specific optimization is essential.
---
## GEO Analysis Criteria (Updated)
### 1. Citability Score (25%)
**Optimal passage length: 134-167 words** for AI citation. And **~44% of AI
citations come from the first 30% of a page** (SE Ranking study), front-load
your most citable, self-contained answer rather than burying it below the fold.
**Strong signals:**
- Clear, quotable sentences with specific facts/statistics
- Self-contained answer blocks (can be extracted without context)
- Direct answer in first 40-60 words of section
- Claims attributed with specific sources
- Definitions following "X is..." or "X refers to..." patterns
- Unique data points not found elsewhere
**Weak signals:**
- Vague, general statements
- Opinion without evidence
- Buried conclusions
- No specific data points
### 2. Structural Readability (20%)
**92% of AI Overview citations come from top-10 ranking pages**, but 47% come from pages ranking below position 5, demonstrating different selection logic.
**Strong signals:**
- Clean H1->H2->H3 heading hierarchy
- Question-based headings (matches query patterns)
- Short paragraphs (2-4 sentences)
- Tables for comparative data
- Ordered/unordered lists for step-by-step or multi-item content
- FAQ sections with clear Q&A format
**Weak signals:**
- Wall of text with no structure
- Inconsistent heading hierarchy
- No lists or tables
- Information buried in paragraphs
### 3. Multi-Modal Content (15%)
Content with multi-modal elements sees **156% higher selection rates**.
**Check for:**
- Text + relevant images
- Video content (embedded or linked)
- Infographics and charts
- Interactive elements (calculators, tools)
- Structured data supporting media
### 4. Authority & Brand Signals (20%)
**Strong signals:**
- Author byline with credentials
- Publication date and last-updated date
- **Recency**, content under 3 months old is ~3x more likely to be cited in AI answers; pages left stale 6+ months lose citation eligibility (SE Ranking, 1.3M-citation study). A scheduled refresh program is one of the highest-leverage GEO plays.
- Citations to primary sources (studies, official docs, data)
- Organization credentials and affiliations
- Expert quotes with attribution
- Entity presence in Wikipedia, Wikidata
- Mentions on Reddit, YouTube, LinkedIn
**Weak signals:**
- Anonymous authorship
- No dates
- No sources cited
- No brand presence across platforms
### 5. Technical Accessibility (20%)
**AI crawlers do NOT execute JavaScript.** Server-side rendering is critical.
**Check for:**
- Server-side rendering (SSR) vs client-only content
- AI crawler access in robots.txt
- llms.txt file presence and configuration
- RSL 1.0 licensing terms
---
## AI Crawler Detection
Check `robots.txt` for these AI crawlers:
| Crawler | Owner | Purpose | Obeys robots.txt? |
|---------|-------|---------|---|
| GPTBot | OpenAI | **Model training only** (NOT ChatGPT Search) | yes |
| OAI-SearchBot | OpenAI | **ChatGPT Search citability** (the crawler that decides it) | yes |
| ChatGPT-User | OpenAI | ChatGPT browsing (user-triggered) | no (user-triggered) |
| ClaudeBot | Anthropic | **Model training only** (NOT Claude's search features) | yes |
| Claude-SearchBot | Anthropic | **Claude/Claude.ai search-result citability** (the crawler that decides it) | yes |
| Claude-User | Anthropic | Claude browsing on a user's behalf (user-triggered) | no (user-triggered) |
| PerplexityBot | Perplexity | Perplexity AI search | yes |
| CCBot | Common Crawl | Training data (often blocked) | yes |
| Bytespider | ByteDance | TikTok/Douyin AI | yes |
| cohere-ai | Cohere | Cohere models | yes |
| Google-Extended | Google | **Gemini/Vertex training & grounding only** (NOT Google Search) | yes |
| Google-CloudVertexBot | Google | Site-owner-requested Vertex AI Agent crawls | yes |
| Google-Agent | Google | Agentic browsing (Project Mariner), acts for a user | **no (user-triggered)** |
| Google-NotebookLM | Google | Fetches individual user-added source URLs | **no (user-triggered)** |
| Google Messages | Google | User-triggered fetch | **no (user-triggered)** |
| Applebot-Extended | Apple | **Apple Intelligence / generative-AI training data opt-out only** (NOT Siri, Spotlight, or Safari search; does not itself crawl, it labels content already fetched by Applebot) | yes |
Sources: [OpenAI crawlers](https://platform.openai.com/docs/bots),
[Google crawlers overview](https://developers.google.com/search/docs/crawling-indexing/overview-google-crawlers),
[Anthropic crawler support article](https://support.anthropic.com/en/articles/8896518-does-anthropic-crawl-data-from-the-web-and-how-can-site-owners-block-the-crawler),
[Apple Applebot-Extended support article](https://support.apple.com/en-us/119829).
Anthropic's current crawler support article documents only ClaudeBot, Claude-User,
and Claude-SearchBot; it does not list `anthropic-ai`, so the previously-unverified
`anthropic-ai` row has been removed rather than kept as a guess.
**Recommendation:** Allow OAI-SearchBot, Claude-SearchBot, and PerplexityBot for AI
search visibility. GPTBot, ClaudeBot, CCBot, and Applebot-Extended are training-only
signals -- allow or block them on licensing preference, not on search-visibility
grounds.
### Check the right bot for the claim you are making
Two pairs are routinely conflated. **Each claim below may only be supported by its own
bot's robots.txt status** -- check them separately and report them separately.
| Claim you want to make | Bot to check | Bot that does NOT support this claim |
|---|---|---|
| "Content is citable in ChatGPT Search" | `OAI-SearchBot` | `GPTBot` |
| "Content is available for OpenAI model training" | `GPTBot` | `OAI-SearchBot` |
| "Content can be used for Gemini/Vertex training & grounding" | `Google-Extended` | `Googlebot` |
| "Content is eligible for Google Search / AI Overviews" | `Googlebot` | `Google-Extended` |
| "Content is citable in Claude's search features" | `Claude-SearchBot` | `ClaudeBot` |
| "Content is available for Anthropic model training" | `ClaudeBot` | `Claude-SearchBot` |
| "Content can be used for Apple Intelligence training" | `Applebot-Extended` | `Applebot` |
| "Content is discoverable via Siri, Spotlight, or Safari search" | `Applebot` | `Applebot-Extended` |
- **`Google-Extended` governs Gemini and Vertex AI training and grounding use only.
It does not affect inclusion in ordinary Google Search, or in AI Overviews and AI
Mode, both of which are served from the `Googlebot` index.** Never score
`Google-Extended` as a "Google Search readiness" signal, and never cite a blocked
`Google-Extended` as evidence that a site is missing from Google Search.
- **`OAI-SearchBot` is the crawler that determines ChatGPT Search citability.
`GPTBot` is OpenAI's separate training crawler.** Checking `GPTBot` access tells
you nothing about whether ChatGPT Search can cite the page. A site that blocks
`GPTBot` and allows `OAI-SearchBot` is fully citable in ChatGPT Search.
- **`Claude-SearchBot` is the crawler that determines citability in Claude's own
search features. `ClaudeBot` is Anthropic's separate training crawler** (per
Anthropic's crawler support article). Checking `ClaudeBot` access tells you
nothing about Claude search citability, and vice versa; report each separately.
- **`Applebot-Extended` is a training-data opt-out signal, not a crawler that
fetches pages itself.** Per Apple's support article, disallowing
`Applebot-Extended` opts a site out of Apple Intelligence / generative-model
training use, but the page remains discoverable through Siri, Spotlight, and
Safari as long as `Applebot` itself is allowed. Never cite a blocked
`Applebot-Extended` as evidence a site is missing from Apple's search surfaces.
Do not use these names interchangeably in report prose. When reporting crawler access,
name the specific user-agent that was checked and the specific capability it governs.
> **User-triggered fetchers ignore robots.txt by design** (Google-Agent, Google-NotebookLM, Google Messages, ChatGPT-User). robots.txt cannot block them, use server-side access controls. Google's canonical crawling/robots reference moved to **developers.google.com/crawling** (migrated 2025-11-20); IP-range files now live at `/crawling/ipranges/` and `googlebot.json` was renamed `common-crawlers.json`. Emerging: **Web Bot Auth** (RFC 9421) lets bots authenticate via a `Signature-Agent` header + key directory (used by Google-Agent); reverse-DNS verification remains the fallback.
---
## llms.txt Standard
Read `references/llmstxt-evidence.md` for the primary-source evidence (Mueller, Illyes, SE Ranking 300k-domain study, OtterlyAI server-log audit) on why `/llms.txt` is not currently a citation lever for major AI search systems. claude-seo reports presence but assigns no citation-ranking weight.
> **Google now states this explicitly.** Google's AI optimization guide, introduced
> 2026-05-15 and clarified 2026-06-15, says `llms.txt` and other AI-text files are
> not needed for Google Search and do not help or hurt visibility or rankings.
> They may still serve non-Google systems. Never recommend `llms.txt` as a Google
> ranking or citation lever. Source:
> developers.google.com/search/docs/fundamentals/ai-optimization-guide
The emerging **llms.txt** standard provides AI crawlers with structured content guidance.
**Location:** `/llms.txt` (root of domain)
**Format:**
```
# Title of site
> Brief description
## Main sections
- [Page title](url): Description
- [Another page](url): Description
## Optional: Key facts
- Fact 1
- Fact 2
```
**Check for:**
- Presence of `/llms.txt`
- Structured content guidance
- Key page highlights
- Contact/authority information
---
## RSL 1.0 (Really Simple Licensing)
New standard (December 2025) for machine-readable AI licensing terms.
**Backed by:** Reddit, Yahoo, Medium, Quora, Cloudflare, Akamai, Creative Commons
**Check for:** RSL implementation and appropriate licensing terms.
---
## Platform-Specific Optimization
| Platform | Key Citation Sources | Optimization Focus |
|----------|---------------------|-------------------|
| **Google AI Overviews** | Strongly ranking-correlated, cites pages that already rank well | Traditional SEO + passage optimization |
| **Google AI Mode** (custom version of Gemini 2.5) | Weakly ranking-correlated; broader pool (~9 domains cited/query, Ahrefs) | Distinct surface: freshness, entity authority, citable passages beyond position 5 |
| **ChatGPT** | Wikipedia (47.9%), Reddit (11.3%) | Entity presence, authoritative sources |
| **Perplexity** | Reddit (46.7%), Wikipedia | Community validation, discussions |
| **Bing Copilot** | Bing index, authoritative sites | Bing SEO, IndexNow |
> **Two Google citation engines, not one.** AI Mode and AI Overviews reach the
> same conclusion ~86% of the time but cite the same URLs only **13.7%** of the
> time (Ahrefs study, 540K query pairs). Treat them as separate surfaces: ranking
> well in classic Search feeds AI Overviews, but AI Mode draws from a broader pool
> where freshness and entity authority outweigh raw position. Score both.
>
> **AI Mode is also a booking surface (2026-08-27).** Flight price tracking
> with email alerts (180+ countries and territories), hotel booking through
> integrated partners, and fares shown in points or miles now happen inside
> AI Mode. Travel and hospitality clients should check partner eligibility;
> nothing here is a documented ranking change.
>
> **UX is now unified, surfaces still distinct.** At Google I/O 2026 (2026-05-19)
> Google merged AI Overviews and AI Mode into "one seamless AI Search experience"
> (question → AI Overview → follow-up in AI Mode) with a new intelligent Search
> box. The *experience* is one flow, but the two citation engines remain
> technically distinct (different models/link sets), keep scoring both.
### Citation surfaces & controls in AI Search (2026)
Google added many AI citation/source surfaces across AI Overviews **and** AI Mode (May 2026):
- **Preferred Sources**, an eligible domain or subdomain can be selected by a
user, making its content more likely to appear in that user's Top Stories and
eligible for a preferred badge in AI Mode or AI Overviews. This is a
**per-user preference**, not a documented general ranking signal. Publishers
may offer Google's interactive button or a deeplink, but should not promise a
site-wide ranking lift. Source:
developers.google.com/search/docs/appearance/preferred-sources
- **"Highly Cited" badges**, earned via original primary reporting that other articles cite.
- **Community Perspectives**, elevates Reddit/forum/firsthand content.
- Inline links, desktop hover **Link Previews**, and prominent link carousels.
**Controlling AI-feature appearance:** there is **no AI-specific opt-out file**. Appearance in AI Overviews and AI Mode is governed by standard preview/index directives, `nosnippet`, `data-nosnippet`, `max-snippet`, `noindex` (distinct from the third-party AI-crawler robots controls above). Source: developers.google.com/search/docs/appearance/ai-features
**Search agents (live, not just WebMCP):** Google's "Information Agents" run in the background to monitor topics, plus agentic booking/calling for select categories (rolling out to US users, summer 2026), so agent-friendly-page optimization (real interactive elements, accessibility tree, layout stability) now matters for actions, not only citations.
---
## Output
Generate `GEO-ANALYSIS.md` with:
1. **GEO Readiness Score: XX/100**
2. **Platform breakdown** (Google AIO, ChatGPT, Perplexity scores)
3. **AI Crawler Access Status** -- report each crawler separately with the
capability it governs. Training access (`GPTBot`, `Google-Extended`, `CCBot`,
`ClaudeBot`, `Applebot-Extended`) and search citability (`OAI-SearchBot`,
`Googlebot`, `PerplexityBot`, `Claude-SearchBot`, `Applebot`) are distinct
findings and must never be merged into one line.
4. **llms.txt Status** (present, missing, recommendations)
5. **Brand Mention Analysis** (presence on Wikipedia, Reddit, YouTube, LinkedIn)
6. **Passage-Level Citability** (optimal 134-167 word blocks identified)
7. **Server-Side Rendering Check** (JavaScript dependency analysis)
8. **Top 5 Highest-Impact Changes**
9. **Schema Recommendations** (for AI discoverability)
10. **Content Reformatting Suggestions** (specific passages to rewrite)
---
## Quick Wins
1. Add "What is [topic]?" definition in first 60 words
2. Create 134-167 word self-contained answer blocks
3. Add question-based H2/H3 headings
4. Include specific statistics with sources
5. Add publication/update dates
6. Implement Person schema for authors
7. Allow key AI crawlers in robots.txt
## Medium Effort
1. Create `/llms.txt` file (optional: ignored by Google Search; may help other AI crawlers)
2. Add author bio with credentials + Wikipedia/LinkedIn links
3. Ensure server-side rendering for key content
4. Build entity presence on Reddit, YouTube
5. Add comparison tables with data
6. Implement FAQ sections (structured, not schema for commercial sites)
## High Impact
1. Create original research/surveys (unique citability)
2. Build Wikipedia presence for brand/key people
3. Establish YouTube channel with content mentions
4. Implement comprehensive entity linking (sameAs across platforms)
5. Develop unique tools or calculators
## DataForSEO Integration (Optional)
If DataForSEO MCP tools are available, use `ai_optimization_chat_gpt_scraper` to check what ChatGPT web search returns for target queries (real GEO visibility check) and `ai_opt_llm_ment_search` with `ai_opt_llm_ment_top_domains` for LLM mention tracking across AI platforms.
## Error Handling
| Scenario | Action |
|----------|--------|
| URL unreachable (DNS failure, connection refused) | Report the error clearly. Do not guess site content. Suggest the user verify the URL and try again. |
| AI crawlers blocked by robots.txt | Report exactly which crawlers are blocked and which are allowed. Provide specific robots.txt directives to add for enabling AI search visibility. |
| No llms.txt found | Note the absence (optional file; Google Search ignores it) and provide a ready-to-use llms.txt template for non-Google AI crawlers. |
| No structured data detected | Report the gap and provide specific schema recommendations (Article, Organization, Person) for improving AI discoverability. |
## FLOW Framework Integration
For prompt-guided AI content optimization, use `/seo flow optimize <url>`, FLOW's 21 optimize-stage prompts complement GEO's citability and structure analysis with evidence-led AI prompts.
Questions
Questions people ask before installing
Which bots should an ecommerce store allow?
The skill’s recommendation: allow OAI-SearchBot, Claude-SearchBot and PerplexityBot for visibility; decide GPTBot, ClaudeBot, CCBot and Applebot-Extended on licensing grounds, because they only govern training.
Does it check whether my brand is mentioned on Reddit or Wikipedia?
It reports brand mention presence as a section of the output, but without DataForSEO it works from what it can fetch and what you tell it. It does not search Reddit itself.
What is the GEO readiness score worth?
It is the plugin author’s weighting of five criteria. Useful to compare two of your own pages; not comparable with anyone else’s tool.