You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
README robots/llms: unlike the docs site, the README already scores 15/robots and 14/llms — GitHub's own repo infrastructure gives it partial credit here.
Platform citability: strong per-engine potential — ChatGPT 50/100, Perplexity 65/100, Google AI 80/100 — driven by FAQ schema, content freshness, and SSR-safe (non-JS-dependent) rendering.
🚨 Critical Gaps
robots.txt missing AI bot rules — 0/18 on the docs site (and every one of the 20 sitemap pages scores 0 here too). No Allow rules for GPTBot, ClaudeBot, or PerplexityBot were found.
/llms.txt missing — 0/18 across the entire docs site and sitemap (all 20 pages average 0.0 on this dimension) — no LLM-oriented site summary exists.
AI discovery endpoints missing — 0/6: no /.well-known/ai.txt, /ai/summary.json, /ai/faq.json, or /ai/service.json on either target.
Schema JSON-LD missing on README — 0/16: no WebSite, Organization, or FAQPage structured data on the GitHub repo homepage.
Hidden text / cloaking detected on the docs homepage (display:none/visibility:hidden content, e.g. "CtrlK") — flagged as a medium-severity negative signal that AI crawlers may penalize.
4 sitemap pages fall into the "Critical" band (/about 35, /blog/7 33, /blog/8 35, /blog/10 35) — mostly driven by weaker schema (3/16) and brand/entity scores (2–4/10) on paginated blog listing pages.
🔧 Recommended Fixes (prioritized by impact)
Create robots.txt with explicit Allow rules for AI crawlers (GPTBot, ClaudeBot, PerplexityBot, Google-Extended) on the docs site — this single change would recover up to 18 points on every one of the 20 audited pages.
Generate /llms.txt via geo llms --base-url https://github.github.com/gh-aw — recovers up to 18 points site-wide and gives AI engines a structured index of the site.
Add WebSite/Organization/FAQPage JSON-LD schema to the README — closes the 16-point schema gap that exists only on the GitHub repo page (the docs site already has this).
Remove/fix hidden text (display:none/visibility:hidden) on the docs homepage to eliminate the -3 negative-signal penalty and cloaking risk.
Add /.well-known/ai.txt, /ai/summary.json, /ai/faq.json, /ai/service.json to define AI crawler permissions and machine-readable site/service summaries (+6 points, "ai_discovery" category currently at 0 on both targets).
Front-load key information in the first 30% of page content for AI snippet selection (docs homepage has_front_loading: false), and add numerical/statistics data (+40% AI visibility per the tool's own estimate).
Improve schema on paginated blog pages (/blog/2–/blog/10) — these average only 5.75/16 on schema vs. 13/16 on the homepage, dragging the sitemap-wide average down to 38.15.
📋 Full Breakdown by Category (docs homepage vs. README)
reacted with thumbs up emoji reacted with thumbs down emoji reacted with laugh emoji reacted with hooray emoji reacted with confused emoji reacted with heart emoji reacted with rocket emoji reacted with eyes emoji
Uh oh!
There was an error while loading. Please reload this page.
GEO Audit Report — github/gh-aw
Audit Date: 2026-08-01
Run: §30708587097
📊 Scores
github.github.com/gh-aw/, homepage)✅ Top Strengths
WebSite,Organization,SoftwareApplication, andFAQPagetypes all detected.lang="en", RSS feed (/blog/rss.xml), and freshness date (2026-05-09) all present.🚨 Critical Gaps
robots.txtmissing AI bot rules — 0/18 on the docs site (and every one of the 20 sitemap pages scores 0 here too). NoAllowrules for GPTBot, ClaudeBot, or PerplexityBot were found./llms.txtmissing — 0/18 across the entire docs site and sitemap (all 20 pages average 0.0 on this dimension) — no LLM-oriented site summary exists./.well-known/ai.txt,/ai/summary.json,/ai/faq.json, or/ai/service.jsonon either target.WebSite,Organization, orFAQPagestructured data on the GitHub repo homepage.display:none/visibility:hiddencontent, e.g. "CtrlK") — flagged as a medium-severity negative signal that AI crawlers may penalize./about35,/blog/733,/blog/835,/blog/1035) — mostly driven by weaker schema (3/16) and brand/entity scores (2–4/10) on paginated blog listing pages.🔧 Recommended Fixes (prioritized by impact)
robots.txtwith explicitAllowrules for AI crawlers (GPTBot, ClaudeBot, PerplexityBot, Google-Extended) on the docs site — this single change would recover up to 18 points on every one of the 20 audited pages./llms.txtviageo llms --base-url https://github.github.com/gh-aw— recovers up to 18 points site-wide and gives AI engines a structured index of the site.WebSite/Organization/FAQPageJSON-LD schema to the README — closes the 16-point schema gap that exists only on the GitHub repo page (the docs site already has this).display:none/visibility:hidden) on the docs homepage to eliminate the -3 negative-signal penalty and cloaking risk./.well-known/ai.txt,/ai/summary.json,/ai/faq.json,/ai/service.jsonto define AI crawler permissions and machine-readable site/service summaries (+6 points, "ai_discovery" category currently at 0 on both targets).has_front_loading: false), and add numerical/statistics data (+40% AI visibility per the tool's own estimate)./blog/2–/blog/10) — these average only 5.75/16 on schema vs. 13/16 on the homepage, dragging the sitemap-wide average down to 38.15.📋 Full Breakdown by Category (docs homepage vs. README)
Sitemap-wide average breakdown (20 pages): robots 0.0, llms 0.0, schema 5.75, meta 14.0, content 10.75, signals 5.4, ai_discovery 0.0, brand_entity 3.2, negative_penalty -0.95.
Trust Stack (docs homepage): composite score 16, grade C, medium trust level — Technical 3/5, Identity 3/5, Social 4/5, Academic 2/5, Consistency 4/5.
Platform citation scores: ChatGPT 50/100, Perplexity 65/100, Google AI 80/100 (citability score 78/100 underlying all three).
📄 Sitemap Page Scores (20 pages audited)
Band distribution: 16 pages "Foundation", 4 pages "Critical", 0 "Good"/"Excellent".
Top 5 pages:
/gh-aw(homepage)/blog/2026-01-13-meet-the-workflows-continuous-improvement/blog/2026-01-13-meet-the-workflows-continuous-refactoring/blog/2026-01-12-welcome-to-pelis-agent-factory/blog/2026-01-13-meet-the-workflows-advanced-analyticsWorst 5 pages:
/blog/7/about/blog/8/blog/10/blog/9All 20 URLs returned HTTP 200 with no failed crawls.
Automated audit powered by geo-optimizer-skill · Run logs
All reactions