AI Search SEO: Citation Strategies skill
Use GEO and AEO as legacy labels for AI citation readiness.
by AgriciDaniel·MIT license·★ 2,219 Stars on the repo·GitHub ↗
Files of AI Search SEO: Citation Strategies
AgriciDaniel/
Show the full text281 lines
AI Search SEO: Citation Strategies
Use GEO and AEO as legacy labels for AI citation readiness. For Google, the Google Search Central AI features guidance is explicit: optimization for AI Overviews and AI Mode is SEO, not a separate discipline.
Core AI Search SEO Research
Princeton GEO Paper (KDD 2024)
The Princeton GEO paper reports that content changes can boost AI visibility by up to 40% in test conditions. Treat the result as research on AI search surfaces, not as a separate Google ranking discipline.
| Technique | Improvement |
|---|---|
| Citing authoritative sources | +115.1% visibility (5th-ranked sites, main experiment) |
| Quotation addition | +28% (main experiment); +37% (Perplexity.ai validation, Table 7) |
| Statistics addition | +41% (main experiment); +22% (Perplexity.ai validation, Table 7) |
| FAQPage entity markup | May aid AI citation as an entity signal for visible Q&A; no Google rich result; impact unverified |
Traditional keyword stuffing performs worse than baseline in generative engines.
Cross-Platform Citation Divergence
- Only 11% of domains are cited by both ChatGPT and Perplexity (Digital Bloom, 2025; domain-level, not URL-level; AI Overviews not included in that study)
- 80% of LLM citations don't rank in Google's top 100 (Ahrefs, Aug 2025) - classic organic rankings alone are a poor predictor of AI citation
- Brands are 6.5x more likely to be cited through third-party sources than their own domains (AirOps, Oct 2025) - earned media dominates AI visibility
Kevin Indig's AI Search Pipeline (Jan 5, 2026)
Three critical stages:
- Retrieval: Which pages enter the candidate set
- Server response time under 200ms TTFB
- Metadata relevance
- Content must be in HTML (not behind JS)
- Citation: Which sources get mentioned
- Content freshness dominates (70%+ cited pages updated within 12 months)
- Content within 3 months performs best
- Trust: Which citations users click
- Brand recognition
- Source authority
Content Format Impact on Citations
| Format | Impact | Source |
|---|---|---|
| Listicles | Often over-index in vendor datasets; exact shares vary and are directional | |
| Tables/structured data | May improve extractability; specific 2.5x vendor claim is unverified | |
| Long-form (2,000+ words) | Often over-indexes in vendor studies; treat multipliers as directional | |
| FAQPage entity markup | May aid AI citation as an entity signal for visible Q&A; no Google rich result; impact unverified | |
| Content with statistics | Original, sourced statistics improve citeability; exact lifts vary by study | |
| Sections of 120-180 words between headings | Self-contained passages likely help extraction; exact lift is unverified | |
Comparison tables with <thead> |
May improve extraction; attributed SEL figure is unverified |
Passage-Level Extractability (2026)
Google's AI systems fragment pages and evaluate self-contained answer passages, not just whole documents. Start each H2 with an approximately 50-word direct-answer sentence, then build a 120-180 word passage that can stand alone if quoted or summarized. Support it with named entities, dates, source attribution, and a specific example.
Entity density and demonstrated first-hand Experience break ties. A single clean passage with original testing, named tools, and verifiable evidence can earn an AI Overview citation even when the full page is not the strongest organic result.
Platform-Specific Citation Patterns
Each AI platform has distinct content preferences:
| Platform | Favored Content Type | Key Bias |
|---|---|---|
| ChatGPT | "Best X" listicles | 43.8% of citations are list-format content |
| Perplexity | Reddit discussions | 6.6% of all citations come from Reddit |
| AI Overviews | Google properties | 23% of citations favor Google-owned sources |
2026 wrinkle: AI Overviews now highlight links from a user's subscribed publications, so publisher subscriptions can influence which sources users see inside the AI answer (Nieman Lab, 2026-05).
Perplexity content decay: Citation relevance begins declining 2-3 days post-publication - Perplexity heavily weights recency, making it the most freshness-dependent platform. Content older than 1 week sees sharp citation drops.
Content Freshness Requirements
- 76.4% of ChatGPT's most-cited pages updated within 30 days (Ahrefs, ~17M citations)
- URLs cited in AI results are 25.7% fresher than traditional search
- Content < 3 months old is 3x more likely to get cited
- Action: Update critical content when facts, screenshots, pricing, methods, source availability, or SERP intent have materially changed
Off-Site Signals (Dominate AI Visibility)
Ahrefs Study (Dec 2025, 75,000 brands)
| Factor | Correlation with AI Visibility |
|---|---|
| YouTube mentions | 0.737 (strongest) |
| Branded web mentions | 0.656-0.709 |
| Domain Rating | 0.266-0.326 |
| Backlinks | 0.218 (dramatically weaker than expected) |
Platform-Specific Citation Rates
YouTube:
- Citations in AI Overviews up 414% (Q1 2025, NP Digital, 10K+ AIO analysis)
- How-to videos up 651%
- Visual demos up 592%
- 200x more cited than any other video platform
- Optimization: keywords in titles/transcripts, Q&A-style, 10+ min, public transcripts
Reddit:
- Citations surged 1.30% → 7.15% (450% growth)
- Google's $60M annual API deal
- 2.2-21% of AI Overview citations by query type
- Strategy: Authentic participation in 3-5 subreddits BEFORE any promotional content
Review Platforms (B2B):
- G2 accounts for 22-23% of review-platform citations (Radix via G2's own blog; self-reported - Hall.com's independent analysis found G2 at only 8.25% of B2B software citations in ChatGPT)
- 33% of review citations come from G2 (Profound via G2's blog; treat as directional)
- Multi-platform presence: 4.6-6.3 citations vs 1.8 without (2.6-3.5x multiplier)
Wikipedia/Wikidata:
- 7.8% of all ChatGPT citations (Profound)
- Used as "credibility tiebreaker" when sources conflict
Budget Allocation
Recommended: 40% owned content / 60% earned media (Most companies allocate 90/10 - this is wrong for AI search SEO)
88-92% of AI citations come from off-site signals in vendor-reported datasets. Treat this as directional, not a universal law.
AI Crawler Technical Requirements
| Crawler | JavaScript Rendering |
|---|---|
| GPTBot (OpenAI) | No |
| ChatGPT-User | No |
| ClaudeBot | No |
| PerplexityBot | No |
| Googlebot | Yes |
Critical: Content behind JavaScript is invisible to ChatGPT, Claude, Perplexity. Use SSR, SSG, or ISR. Test by disabling JS and reloading.
Google's Official Gen-AI Guidance
Google's stance holds: optimization for AI Overviews and AI Mode is SEO. There is no special schema for gen-AI features, and Google does not need llms.txt. Use standard crawlable HTML, Article schema with author and Organization entities, helpful content, clear source attribution, and fast server responses.
AI Crawler Traffic Growth
- Cloudflare AI crawling rose 32% YoY across all monitored sites
- GPTBot traffic grew 305% YoY
- PerplexityBot traffic grew 157,490% YoY (from near-zero baseline)
- 65% of AI bot hits target content published within the past year (Seer Interactive)
- freshness is a retrieval signal, not just a citation signal
Performance Requirements for AI Retrieval
- Server response time under 200ms TTFB (Kevin Indig pipeline)
- TTFB above 600ms may reduce crawl or extraction reliability
- Some crawlers and retrieval systems use short practical timeouts; verify per crawler
- Core Web Vitals are a constraint, not a growth lever - good CWV doesn't reliably outperform, but severe LCP failure creates disadvantage (Search Engine Land, 107,352 pages)
- Top 10 domains capture 46% of all ChatGPT citations per topic (Growth Memo, Mar 2026)
- Slow pages may miss crawl, fetch, or extraction opportunities
- Vercel analysis of 500+ million GPTBot fetches found zero evidence of JS execution
robots.txt for AI Visibility
User-agent: GPTBot
Allow: /
User-agent: ChatGPT-User
Allow: /
User-agent: ClaudeBot
Allow: /
User-agent: PerplexityBot
Allow: /
llms.txt Standard
Google does not need llms.txt for AI Overviews or AI Mode. Treat the file as an optional site inventory for non-Google tools, not a ranking or citation lever. Do not spend AI search SEO budget on llms.txt before crawlability, passage extraction, Article schema, source quality, and entity consistency.
Attribution Gaps
Perplexity visits ~10 pages per query but cites only 3-4. Not all AI responses include citations - optimizing for retrieval is critical. Content must enter the candidate set before citation is possible.
AI Search Case Study Results
These examples are illustrative, vendor-reported, and not independently verified. Do not reuse the numbers as factual benchmarks without primary confirmation.
| Company | Results | Timeframe |
|---|---|---|
| Go Fish Digital | +43% AI traffic, +83% conversions, 25x conversion rate | 3 months |
| Netpeak USA | +120% revenue, +693% AI visits | Ongoing |
| Nine Peaks Media | 36% visibility improvement, first ChatGPT citations | Ongoing |
| ABM Agency/Chemours | 82% ChatGPT mention rate, $90M+ pipeline | Ongoing |
| Smart Rent | 32% SQL increase, 40% faster pipeline | Ongoing |
Entity-First SEO
Every page should unambiguously represent ONE canonical entity. Google Knowledge Graph: 800B facts about 8B entities.
Entity building timeline (3-6 months):
- Create entity map with Wikidata Q-IDs
- Establish Wikipedia/Wikidata presence only when the entity meets independent notability requirements; disclose conflicts of interest and follow platform policy
- Build entity consistency across all platforms (exact same name)
- Practice "controlled co-occurrence" via third-party mentions
- Earn external citations from recognized publications
Readability and AI Search Connection
Readability can support AI citation rates, but the lift is not independently verified. Use Flesch 60-75 as a clarity heuristic, not as a citation guarantee.
Flesch Score & AI Citation Rates
- Commercial platform reports associate Flesch 60-75 with more AI citations; this is vendor-reported internal data with no independent verification.
- Teams improving Flesch from 52→68 saw parallel citation lifts within two crawl windows
- Content that is too complex (Flesch <50) or too simple (Flesch >80) gets fewer citations - AI systems prefer fluent, authoritative writing
Citation Position Bias
- 44.2% of all LLM citations come from the first 30% of text is a single-source Growth Memo finding (Feb 2026, Kevin Indig). Treat it as directional support for answer-first formatting, not a settled benchmark.
- Direct answers in the first 1-2 sentences of each section maximize extractability for AI systems
AI Search Tactic Combinations
Princeton GEO paper (KDD 2024) findings on readability-related tactics:
- Fluency optimization = 15-30% visibility boost
- Statistics addition = up to 41% visibility boost
- Fluency + Statistics combined outperforms any single tactic by 5.5%
- Keyword stuffing performs -10% WORSE than baseline
FLOW evidence triple is mandatory for AI-citation readiness. AI assistants extract claims that have year anchor in prose, inline publisher + title, and URL with retrieval date. Stats without the triple are less likely to surface in citations. See flow-alignment.md.
Schema & Structure for AI Citation
- Comparison tables with proper HTML (
<thead>,<tbody>) = 47% higher AI citation rates (attributed to SEL; primary source unlocatable - treat as directional) - Structured data helps machine understanding, but do not claim all major AI platforms use schema during citation selection without current primary evidence
Platform-Specific Citation Behaviors
| Platform | Key Behavior | Readability Preference |
|---|---|---|
| ChatGPT | Wikipedia = 7.8% of citations; SearchGPT: 87% match Bing top 10 | Prefers well-structured, fluent content |
| Perplexity | Reddit = 46.7% of top-10 sources; strongest depth correlation (0.191) | 2-3 day content decay; heavily weights recency |
| AI Overviews | 93.67% from top-10 organic; avg 10.2 links per response | Prefers established authority + clear answers |
Only 11% of domains are cited by both ChatGPT and Perplexity (Digital Bloom). Only 12% of URLs cited by ChatGPT, Perplexity, and Copilot rank in Google's top 10 (Ahrefs).
Content Freshness for AI Citation
- 65% of AI bot hits target content published within the past year (Seer Interactive)
- 85% of AI Overview citations come from content <2 years old
- 44% of AI Overview citations come from 2025 content specifically
- 50% of Perplexity citations come from 2025 alone
- Content older than 3 months sees 3x fewer citations
| 1 | # AI Search SEO: Citation Strategies |
| 2 | |
| 3 | Use GEO and AEO as legacy labels for AI citation readiness. For Google, the |
| 4 | Google Search Central AI features guidance is explicit: optimization for AI |
| 5 | Overviews and AI Mode is SEO, not a separate discipline. |
| 6 | |
| 7 | ## Core AI Search SEO Research |
| 8 | |
| 9 | ### Princeton GEO Paper (KDD 2024) |
| 10 | The Princeton GEO paper reports that content changes can boost AI visibility |
| 11 | by up to 40% in test conditions. Treat the result as research on AI search |
| 12 | surfaces, not as a separate Google ranking discipline. |
| 13 | |
| 14 | | Technique | Improvement | |
| 15 | |-----------|-------------| |
| 16 | | Citing authoritative sources | +115.1% visibility (5th-ranked sites, main experiment) | |
| 17 | | Quotation addition | +28% (main experiment); +37% (Perplexity.ai validation, Table 7) | |
| 18 | | Statistics addition | +41% (main experiment); +22% (Perplexity.ai validation, Table 7) | |
| 19 | | FAQPage entity markup | May aid AI citation as an entity signal for visible Q&A; no Google rich result; impact unverified | |
| 20 | |
| 21 | Traditional keyword stuffing performs **worse than baseline** in generative engines. |
| 22 | |
| 23 | ### Cross-Platform Citation Divergence |
| 24 | |
| 25 | Only 11% of domains are cited by both ChatGPT and Perplexity (Digital Bloom, 2025; |
| 26 | domain-level, not URL-level; AI Overviews not included in that study) |
| 27 | 80% of LLM citations don't rank in Google's top 100 (Ahrefs, Aug 2025) - classic |
| 28 | organic rankings alone are a poor predictor of AI citation |
| 29 | Brands are 6.5x more likely to be cited through third-party sources than their own |
| 30 | domains (AirOps, Oct 2025) - earned media dominates AI visibility |
| 31 | |
| 32 | ### Kevin Indig's AI Search Pipeline (Jan 5, 2026) |
| 33 | Three critical stages: |
| 34 | |
| 35 | **Retrieval**: Which pages enter the candidate set |
| 36 | Server response time under 200ms TTFB |
| 37 | Metadata relevance |
| 38 | Content must be in HTML (not behind JS) |
| 39 | **Citation**: Which sources get mentioned |
| 40 | Content freshness dominates (70%+ cited pages updated within 12 months) |
| 41 | Content within 3 months performs best |
| 42 | **Trust**: Which citations users click |
| 43 | Brand recognition |
| 44 | Source authority |
| 45 | |
| 46 | ## Content Format Impact on Citations |
| 47 | |
| 48 | | Format | Impact | Source | |
| 49 | |--------|--------|--------| |
| 50 | | Listicles | Often over-index in vendor datasets; exact shares vary and are directional | |
| 51 | | Tables/structured data | May improve extractability; specific 2.5x vendor claim is unverified | |
| 52 | | Long-form (2,000+ words) | Often over-indexes in vendor studies; treat multipliers as directional | |
| 53 | | FAQPage entity markup | May aid AI citation as an entity signal for visible Q&A; no Google rich result; impact unverified | |
| 54 | | Content with statistics | Original, sourced statistics improve citeability; exact lifts vary by study | |
| 55 | | Sections of 120-180 words between headings | Self-contained passages likely help extraction; exact lift is unverified | |
| 56 | | Comparison tables with `<thead>` | May improve extraction; attributed SEL figure is unverified | |
| 57 | |
| 58 | ### Passage-Level Extractability (2026) |
| 59 | |
| 60 | Google's AI systems fragment pages and evaluate self-contained answer passages, |
| 61 | not just whole documents. Start each H2 with an approximately 50-word |
| 62 | direct-answer sentence, then build a 120-180 word passage that can stand alone |
| 63 | if quoted or summarized. Support it with named entities, dates, source |
| 64 | attribution, and a specific example. |
| 65 | |
| 66 | Entity density and demonstrated first-hand Experience break ties. A single |
| 67 | clean passage with original testing, named tools, and verifiable evidence can |
| 68 | earn an AI Overview citation even when the full page is not the strongest |
| 69 | organic result. |
| 70 | |
| 71 | ## Platform-Specific Citation Patterns |
| 72 | |
| 73 | Each AI platform has distinct content preferences: |
| 74 | |
| 75 | | Platform | Favored Content Type | Key Bias | |
| 76 | |----------|---------------------|----------| |
| 77 | | ChatGPT | "Best X" listicles | 43.8% of citations are list-format content | |
| 78 | | Perplexity | Reddit discussions | 6.6% of all citations come from Reddit | |
| 79 | | AI Overviews | Google properties | 23% of citations favor Google-owned sources | |
| 80 | |
| 81 | 2026 wrinkle: AI Overviews now highlight links from a user's subscribed |
| 82 | publications, so publisher subscriptions can influence which sources users see |
| 83 | inside the AI answer (Nieman Lab, 2026-05). |
| 84 | |
| 85 | **Perplexity content decay**: Citation relevance begins declining 2-3 days |
| 86 | post-publication - Perplexity heavily weights recency, making it the most |
| 87 | freshness-dependent platform. Content older than 1 week sees sharp citation drops. |
| 88 | |
| 89 | ## Content Freshness Requirements |
| 90 | |
| 91 | 76.4% of ChatGPT's most-cited pages updated within 30 days (Ahrefs, ~17M citations) |
| 92 | URLs cited in AI results are 25.7% fresher than traditional search |
| 93 | Content < 3 months old is 3x more likely to get cited |
| 94 | **Action**: Update critical content when facts, screenshots, pricing, methods, |
| 95 | source availability, or SERP intent have materially changed |
| 96 | |
| 97 | ## Off-Site Signals (Dominate AI Visibility) |
| 98 | |
| 99 | ### Ahrefs Study (Dec 2025, 75,000 brands) |
| 100 | |
| 101 | | Factor | Correlation with AI Visibility | |
| 102 | |--------|-------------------------------| |
| 103 | | YouTube mentions | 0.737 (strongest) | |
| 104 | | Branded web mentions | 0.656-0.709 | |
| 105 | | Domain Rating | 0.266-0.326 | |
| 106 | | Backlinks | 0.218 (dramatically weaker than expected) | |
| 107 | |
| 108 | ### Platform-Specific Citation Rates |
| 109 | |
| 110 | **YouTube**: |
| 111 | Citations in AI Overviews up 414% (Q1 2025, NP Digital, 10K+ AIO analysis) |
| 112 | How-to videos up 651% |
| 113 | Visual demos up 592% |
| 114 | 200x more cited than any other video platform |
| 115 | Optimization: keywords in titles/transcripts, Q&A-style, 10+ min, public transcripts |
| 116 | |
| 117 | **Reddit**: |
| 118 | Citations surged 1.30% → 7.15% (450% growth) |
| 119 | Google's $60M annual API deal |
| 120 | 2.2-21% of AI Overview citations by query type |
| 121 | Strategy: Authentic participation in 3-5 subreddits BEFORE any promotional content |
| 122 | |
| 123 | **Review Platforms (B2B)**: |
| 124 | G2 accounts for 22-23% of review-platform citations (Radix via G2's own blog; |
| 125 | self-reported - Hall.com's independent analysis found G2 at only 8.25% of B2B |
| 126 | software citations in ChatGPT) |
| 127 | 33% of review citations come from G2 (Profound via G2's blog; treat as directional) |
| 128 | Multi-platform presence: 4.6-6.3 citations vs 1.8 without (2.6-3.5x multiplier) |
| 129 | |
| 130 | **Wikipedia/Wikidata**: |
| 131 | 7.8% of all ChatGPT citations (Profound) |
| 132 | Used as "credibility tiebreaker" when sources conflict |
| 133 | |
| 134 | ### Budget Allocation |
| 135 | Recommended: **40% owned content / 60% earned media** |
| 136 | (Most companies allocate 90/10 - this is wrong for AI search SEO) |
| 137 | |
| 138 | 88-92% of AI citations come from off-site signals in vendor-reported datasets. |
| 139 | Treat this as directional, not a universal law. |
| 140 | |
| 141 | ## AI Crawler Technical Requirements |
| 142 | |
| 143 | | Crawler | JavaScript Rendering | |
| 144 | |---------|---------------------| |
| 145 | | GPTBot (OpenAI) | No | |
| 146 | | ChatGPT-User | No | |
| 147 | | ClaudeBot | No | |
| 148 | | PerplexityBot | No | |
| 149 | | Googlebot | Yes | |
| 150 | |
| 151 | **Critical**: Content behind JavaScript is invisible to ChatGPT, Claude, Perplexity. |
| 152 | Use SSR, SSG, or ISR. Test by disabling JS and reloading. |
| 153 | |
| 154 | ### Google's Official Gen-AI Guidance |
| 155 | |
| 156 | Google's stance holds: optimization for AI Overviews and AI Mode is SEO. There |
| 157 | is no special schema for gen-AI features, and Google does not need llms.txt. |
| 158 | Use standard crawlable HTML, Article schema with author and Organization |
| 159 | entities, helpful content, clear source attribution, and fast server responses. |
| 160 | |
| 161 | ### AI Crawler Traffic Growth |
| 162 | |
| 163 | Cloudflare AI crawling rose 32% YoY across all monitored sites |
| 164 | GPTBot traffic grew 305% YoY |
| 165 | PerplexityBot traffic grew 157,490% YoY (from near-zero baseline) |
| 166 | 65% of AI bot hits target content published within the past year (Seer Interactive) |
| 167 | freshness is a retrieval signal, not just a citation signal |
| 168 | |
| 169 | ### Performance Requirements for AI Retrieval |
| 170 | Server response time under 200ms TTFB (Kevin Indig pipeline) |
| 171 | TTFB above 600ms may reduce crawl or extraction reliability |
| 172 | Some crawlers and retrieval systems use short practical timeouts; verify per crawler |
| 173 | Core Web Vitals are a constraint, not a growth lever - good CWV doesn't reliably |
| 174 | outperform, but severe LCP failure creates disadvantage (Search Engine Land, 107,352 pages) |
| 175 | Top 10 domains capture 46% of all ChatGPT citations per topic (Growth Memo, Mar 2026) |
| 176 | Slow pages may miss crawl, fetch, or extraction opportunities |
| 177 | Vercel analysis of 500+ million GPTBot fetches found zero evidence of JS execution |
| 178 | |
| 179 | ### robots.txt for AI Visibility |
| 180 | |
| 181 | User-agent: GPTBot |
| 182 | Allow: / |
| 183 | User-agent: ChatGPT-User |
| 184 | Allow: / |
| 185 | User-agent: ClaudeBot |
| 186 | Allow: / |
| 187 | User-agent: PerplexityBot |
| 188 | Allow: / |
| 189 | |
| 190 | |
| 191 | ### llms.txt Standard |
| 192 | Google does not need llms.txt for AI Overviews or AI Mode. Treat the file as an |
| 193 | optional site inventory for non-Google tools, not a ranking or citation lever. |
| 194 | Do not spend AI search SEO budget on llms.txt before crawlability, passage extraction, |
| 195 | Article schema, source quality, and entity consistency. |
| 196 | |
| 197 | ## Attribution Gaps |
| 198 | |
| 199 | Perplexity visits ~10 pages per query but cites only 3-4. Not all AI responses |
| 200 | include citations - optimizing for retrieval is critical. Content must enter the |
| 201 | candidate set before citation is possible. |
| 202 | |
| 203 | ## AI Search Case Study Results |
| 204 | |
| 205 | These examples are illustrative, vendor-reported, and not independently |
| 206 | verified. Do not reuse the numbers as factual benchmarks without primary |
| 207 | confirmation. |
| 208 | |
| 209 | | Company | Results | Timeframe | |
| 210 | |---------|---------|-----------| |
| 211 | | Go Fish Digital | +43% AI traffic, +83% conversions, 25x conversion rate | 3 months | |
| 212 | | Netpeak USA | +120% revenue, +693% AI visits | Ongoing | |
| 213 | | Nine Peaks Media | 36% visibility improvement, first ChatGPT citations | Ongoing | |
| 214 | | ABM Agency/Chemours | 82% ChatGPT mention rate, $90M+ pipeline | Ongoing | |
| 215 | | Smart Rent | 32% SQL increase, 40% faster pipeline | Ongoing | |
| 216 | |
| 217 | ## Entity-First SEO |
| 218 | |
| 219 | Every page should unambiguously represent ONE canonical entity. |
| 220 | Google Knowledge Graph: 800B facts about 8B entities. |
| 221 | |
| 222 | Entity building timeline (3-6 months): |
| 223 | Create entity map with Wikidata Q-IDs |
| 224 | Establish Wikipedia/Wikidata presence only when the entity meets independent |
| 225 | notability requirements; disclose conflicts of interest and follow platform policy |
| 226 | Build entity consistency across all platforms (exact same name) |
| 227 | Practice "controlled co-occurrence" via third-party mentions |
| 228 | Earn external citations from recognized publications |
| 229 | |
| 230 | ## Readability and AI Search Connection |
| 231 | |
| 232 | Readability can support AI citation rates, but the lift is not independently |
| 233 | verified. Use Flesch 60-75 as a clarity heuristic, not as a citation guarantee. |
| 234 | |
| 235 | ### Flesch Score & AI Citation Rates |
| 236 | Commercial platform reports associate Flesch 60-75 with more AI citations; |
| 237 | this is vendor-reported internal data with no independent verification. |
| 238 | Teams improving Flesch from 52→68 saw parallel citation lifts within two |
| 239 | crawl windows |
| 240 | Content that is too complex (Flesch <50) or too simple (Flesch >80) gets |
| 241 | fewer citations - AI systems prefer fluent, authoritative writing |
| 242 | |
| 243 | ### Citation Position Bias |
| 244 | **44.2% of all LLM citations come from the first 30% of text** is a |
| 245 | single-source Growth Memo finding (Feb 2026, Kevin Indig). Treat it as |
| 246 | directional support for answer-first formatting, not a settled benchmark. |
| 247 | Direct answers in the first 1-2 sentences of each section maximize |
| 248 | extractability for AI systems |
| 249 | |
| 250 | ### AI Search Tactic Combinations |
| 251 | Princeton GEO paper (KDD 2024) findings on readability-related tactics: |
| 252 | **Fluency optimization** = 15-30% visibility boost |
| 253 | **Statistics addition** = up to 41% visibility boost |
| 254 | **Fluency + Statistics combined** outperforms any single tactic by 5.5% |
| 255 | Keyword stuffing performs -10% WORSE than baseline |
| 256 | |
| 257 | **FLOW evidence triple is mandatory for AI-citation readiness.** AI assistants extract claims that have year anchor in prose, inline publisher + title, and URL with retrieval date. Stats without the triple are less likely to surface in citations. See `flow-alignment.md`. |
| 258 | |
| 259 | ### Schema & Structure for AI Citation |
| 260 | Comparison tables with proper HTML (`<thead>`, `<tbody>`) = **47% higher** |
| 261 | AI citation rates (attributed to SEL; primary source unlocatable - treat as directional) |
| 262 | Structured data helps machine understanding, but do not claim all major AI |
| 263 | platforms use schema during citation selection without current primary evidence |
| 264 | |
| 265 | ### Platform-Specific Citation Behaviors |
| 266 | | Platform | Key Behavior | Readability Preference | |
| 267 | |----------|-------------|----------------------| |
| 268 | | ChatGPT | Wikipedia = 7.8% of citations; SearchGPT: 87% match Bing top 10 | Prefers well-structured, fluent content | |
| 269 | | Perplexity | Reddit = 46.7% of top-10 sources; strongest depth correlation (0.191) | 2-3 day content decay; heavily weights recency | |
| 270 | | AI Overviews | 93.67% from top-10 organic; avg 10.2 links per response | Prefers established authority + clear answers | |
| 271 | |
| 272 | Only 11% of domains are cited by both ChatGPT and Perplexity (Digital Bloom). Only 12% |
| 273 | of URLs cited by ChatGPT, Perplexity, and Copilot rank in Google's top 10 (Ahrefs). |
| 274 | |
| 275 | ### Content Freshness for AI Citation |
| 276 | **65%** of AI bot hits target content published within the past year (Seer Interactive) |
| 277 | **85%** of AI Overview citations come from content <2 years old |
| 278 | **44%** of AI Overview citations come from 2025 content specifically |
| 279 | **50%** of Perplexity citations come from 2025 alone |
| 280 | Content older than 3 months sees 3x fewer citations |
| 281 |
Discussion
Alternatives
Browse more free Claude skills or everything in Marketing.