81/100
Grade B · Ready for agents
Ahead of 39% of 75 sites scanned
Top fixes — ranked by impact
- 01Allow OAI-SearchBot, PerplexityBot, YouBot, Amazonbot, ChatGPT-User, Perplexity-User in robots.txt so AI assistants can read and cite these pages.
- 02To raise content value: add data tables, datasets, or structured records — the most citable content for AI; structure content as questions, lists, and clear definitions; reduce link density and vague "click here" phrasing.
- 03Add an /llms.txt file summarizing your site for AI agents.
- llms.txt No llms.txt found. An llms.txt file gives AI agents a curated summary and key links, guiding them to your most important content. fail +3 pts
-
AI bot access
Some AI search/answer crawlers are blocked: OAI-SearchBot, PerplexityBot, YouBot, Amazonbot, ChatGPT-User, Perplexity-User.
AI assistants can only read, cite, or recommend your site if their crawlers aren't blocked in robots.txt. Search & answer crawlers are what get you cited; training crawlers only affect model corpora.
warn +7 pts
Details
- OAI-SearchBot (AI search) blocked
- PerplexityBot (AI search) blocked
- Claude-SearchBot (AI search) allowed
- DuckAssistBot (AI search) allowed
- YouBot (AI search) blocked
- Amazonbot (AI search) blocked
- Applebot (AI search) allowed
- ChatGPT-User (user-triggered) blocked
- Perplexity-User (user-triggered) blocked
- Claude-User (user-triggered) allowed
- meta-externalfetcher (user-triggered) allowed
- GPTBot (model training) blocked
- ClaudeBot (model training) blocked
- anthropic-ai (model training) blocked
- Google-Extended (model training) blocked
- Applebot-Extended (model training) blocked
- CCBot (model training) blocked
- Bytespider (model training) blocked
- meta-externalagent (model training) blocked
- cohere-ai (model training) blocked
- PetalBot (model training) blocked
-
Content value
Content value: 52% — richest page: /video/docs.
Whether an AI actually uses and cites a page depends on its content: structured, tabular, and data-rich content is the most citable, followed by fact-rich, answer-shaped, cleanly chunked text — plain prose is the least valuable.
warn +5 pts
Details
- Data richness: 0 data table(s), 0 data link(s)
- Substance: 1442 words of main content
- Fact density (numbers, dates, entities)
- Answer-shaped structure (Q-headings, lists, definitions)
- Chunkability (heading hierarchy)
- Self-containedness (low link-density, no vague CTAs)
-
Content quality
Title, description, and image alt text: 67% ideal.
Concise titles, meta descriptions of the right length, and alt text on images give agents accurate, quotable summaries of each page and its media.
warn +2 pts
Details
- Title 116 chars
- Description 126 chars
- Image alt text: 75/147 descriptive · 72 missing
- Machine-readable data No obvious feeds or public APIs detected. Feeds and public APIs let agents consume your data structurally instead of scraping HTML. warn +1 pts
-
Content negotiation
No agent-friendly Markdown representation offered via content negotiation.
Serving a Markdown representation via Accept negotiation gives AI agents clean content without parsing HTML — an emerging, forward-looking signal.
warn
Details
- Serves Markdown for Accept: text/markdown
- Sets Vary: Accept
- Returns 406 for unsupported types (optional per RFC 9110)
- Honors q-values
-
Meta & semantics
5/6 meta/semantic signals present.
Titles, descriptions, canonical URLs, and a clean heading structure help agents identify, summarize, and attribute your pages correctly.
pass +1 pts
Details
- <title>
- meta description
- canonical link
- Open Graph tags
- html lang attribute
- exactly one <h1>
- Content in HTML 2360 words of server-rendered text in semantic <main>/<article>. Most AI crawlers don't execute JavaScript — content must be in the initial HTML or agents simply won't see it. pass
- Structured data (JSON-LD) Valid structured data: WebPage, NewsMediaOrganization. Schema.org JSON-LD tells agents exactly what your entities are (products, articles, your organization) so they can extract facts reliably instead of guessing. Incomplete or invalid markup — missing @context, or a Product with no price — is skipped just like missing markup. pass
-
Sitemaps
Sitemap present with 5 entries.
A sitemap gives crawlers and agents a complete, machine-readable map of your pages so none are missed.
pass
Details
- 5 URLs listed
- Bot reachability Site responds 200 to a bot user-agent. If your WAF or bot filter blocks crawler user-agents, agents can't fetch the page at all — regardless of how well it's built. pass
-
HTTPS & redirects
Served over HTTPS.
HTTPS is a baseline trust and ranking signal; if the insecure http:// version also serves content without redirecting, agents can index a duplicate, non-canonical copy.
pass
Details
- HTTPS served
- Indexing directives No indexing or AI opt-out directives block crawlers. Meta robots tags and X-Robots-Tag headers can quietly tell search and AI crawlers to skip your page — often left on by accident from a staging setup. pass
-
Loading speed
Real-user server response (TTFB) is 540 ms at the 75th percentile.
Crawlers apply tight timeouts and crawl budgets — a slow server response gets fetched less often and sometimes abandoned mid-load. Real-user field data (Chrome UX Report) reflects actual server speed better than a single synthetic ping.
pass
Details
- TTFB 540 ms (real-user)
- LCP 1.4 s
- INP 131 ms
- CLS 0.03