# AI readiness of bbc.com

> Measured on the home page. Overall grade: Poor.

Score 48/100, grade Poor. Measured 2026-09-04T15:06:06.029+00:00, engine ax-audit@3.6.0.

## Checks

### LLMs.txt — 0/100

- **FAIL** /llms.txt not found
  HTTP 404
  Create a /llms.txt file at your site root following the llmstxt.org specification. It should be a Markdown file starting with "# Your Site Name" and include a description, sections, and links.
  https://lucioduran.com/projects/ax-audit/guides/llms-txt#not-found

### Robots.txt — 30/100

- **PASS** /robots.txt exists
- **WARN** 7/8 core AI crawlers configured
  Missing: Claude-SearchBot
  Add explicit User-agent entries for the missing crawlers with Allow: / for each one.
  https://lucioduran.com/projects/ax-audit/guides/robots-txt#missing-crawlers
- **WARN** 22 AI crawler(s) explicitly blocked
  Amazonbot, CCBot, omgili, omgilibot, Claude-Web, ClaudeBot, anthropic-ai, cohere-ai, Bytespider, PetalBot, Applebot-Extended, GPTBot, ChatGPT-User, Google-Extended, PerplexityBot, Perplexity-User, Google-CloudVertexBot, meta-externalagent, OAI-SearchBot, Diffbot, YouBot, FirecrawlAgent
  These crawlers have "Disallow: /" rules. If you want AI agents to access your site, change to "Allow: /" for each blocked crawler.
  https://lucioduran.com/projects/ax-audit/guides/robots-txt#explicitly-blocked
- **PASS** Sitemap directive present
- **WARN** No Content-Signal directive found (optional)
  Declare how crawlers may use your content after access with the Content Signals Policy, e.g.: Content-Signal: search=yes, ai-train=no. Known signals: search, ai-input, ai-train. Generate yours at contentsignals.org.
  https://lucioduran.com/projects/ax-audit/guides/robots-txt#missing-content-signals
- **PASS** 22/48 known AI crawlers have explicit rules

### Agent Card (A2A) — 0/100

- **FAIL** /.well-known/agent.json not found
  HTTP 404
  Create a /.well-known/agent.json file following the A2A (Agent-to-Agent) protocol. It should include name, description, url, and skills fields describing your site's capabilities.
  https://lucioduran.com/projects/ax-audit/guides/agent-json#not-found

### Security.txt — 100/100

- **PASS** /.well-known/security.txt exists
- **PASS** Required field "Contact" present
- **PASS** Required field "Expires" present
- **PASS** Expires date is in the future (2038-01-19)
- **PASS** 2/5 optional fields present

### Meta Tags — 46/100

- **WARN** No AI meta tags (ai:*) found
  Add AI meta tags to your HTML <head>: <meta name="ai:summary" content="Brief description">, <meta name="ai:content_type" content="website">, <meta name="ai:author" content="Your Name">.
  https://lucioduran.com/projects/ax-audit/guides/meta-tags#no-ai-meta
- **WARN** No rel="alternate" link to llms.txt in HTML
  Add to your <head>: <link rel="alternate" type="text/plain" href="/llms.txt" title="LLM-optimized content">
  https://lucioduran.com/projects/ax-audit/guides/meta-tags#no-llms-alternate
- **WARN** No rel="alternate" link to agent.json in HTML
  Add to your <head>: <link rel="alternate" type="application/json" href="/.well-known/agent.json" title="Agent Card">
  https://lucioduran.com/projects/ax-audit/guides/meta-tags#no-agent-alternate
- **WARN** No rel="me" identity links found
  Add rel="me" links to verify your identity across platforms: <link rel="me" href="https://github.com/yourname">, <link rel="me" href="https://twitter.com/yourname">.
  https://lucioduran.com/projects/ax-audit/guides/meta-tags#no-rel-me
- **PASS** OpenGraph required tags present (og:title, og:description, og:url, og:type)
- **WARN** OpenGraph recommended tags missing: og:image, og:site_name
  Add these for richer previews: og:image, og:site_name.
  https://lucioduran.com/projects/ax-audit/guides/meta-tags#og-recommended-missing
- **WARN** Twitter Card required tags missing: twitter:card
  Add these meta tags: <meta name="twitter:card" content="...">.
  https://lucioduran.com/projects/ax-audit/guides/meta-tags#twitter-required-missing

### OpenAPI Spec — 0/100

- **FAIL** /.well-known/openapi.json not found
  HTTP 404
  Create a /.well-known/openapi.json file with your API specification following the OpenAPI 3.x standard. See https://swagger.io/specification/ for the spec.
  https://lucioduran.com/projects/ax-audit/guides/openapi#not-found

### MCP (Model Context Protocol) — 0/100

- **FAIL** /.well-known/mcp.json not found
  HTTP 404
  Create a /.well-known/mcp.json file describing your MCP server configuration. Include name, description, tools, and version fields. See https://modelcontextprotocol.io for the spec.
  https://lucioduran.com/projects/ax-audit/guides/mcp#not-found

### Sitemap — 95/100

- **PASS** Sitemap located: https://www.bbc.com/afrique/sitemap.xml
- **PASS** Content-Type is XML (application/xml)
- **PASS** 100 URL(s) declared
- **PASS** 100% of URLs have <lastmod>
- **WARN** Newest <lastmod> is 4363 days old
  Refresh <lastmod> on URLs that have changed. A stale sitemap signals to crawlers that nothing on the site has updated.
  https://lucioduran.com/projects/ax-audit/guides/sitemap#stale

### TLS / HTTPS — 90/100

- **PASS** Site is served over HTTPS
- **PASS** HTTP requests redirect to HTTPS
- **PASS** HSTS max-age=31536000
- **WARN** HSTS does not include subdomains
  Add includeSubDomains to apply HSTS across api., docs., etc. Required for preload list submission.
  https://lucioduran.com/projects/ax-audit/guides/tls-https#hsts-no-subdomains
- **WARN** HSTS has preload directive but does not satisfy preload-list requirements
  Preload requires max-age >= 31536000 and includeSubDomains. See https://hstspreload.org.
  https://lucioduran.com/projects/ax-audit/guides/tls-https#hsts-preload-invalid

### AI Well-Known — 0/100

- **WARN** 0/5 emerging AI discovery files published
- **WARN** ai.txt not found
  Tried: /.well-known/ai.txt
  Publish /.well-known/ai.txt declaring opt-in/opt-out signals for AI training. See https://site.spawning.ai/spawning-ai-txt for the format.
  https://lucioduran.com/projects/ax-audit/guides/well-known-ai#ai-txt
- **WARN** genai.txt not found
  Tried: /.well-known/genai.txt
  Publish /.well-known/genai.txt declaring your generative-AI usage policy.
  https://lucioduran.com/projects/ax-audit/guides/well-known-ai#genai-txt
- **WARN** ai-plugin.json not found
  Tried: /.well-known/ai-plugin.json, /ai-plugin.json
  Publish /ai-plugin.json (legacy ChatGPT plugin manifest). Schema: name_for_model, description_for_model, api.url. Still consumed by some agents.
  https://lucioduran.com/projects/ax-audit/guides/well-known-ai#ai-plugin
- **WARN** agents.json not found
  Tried: /agents.json, /.well-known/agents.json
  Publish /agents.json describing your site as a callable agent (OpenAgents / Wildcard emerging spec). Includes name, description, and operations[].
  https://lucioduran.com/projects/ax-audit/guides/well-known-ai#agents-json
- **WARN** nlweb.json not found
  Tried: /.well-known/nlweb.json, /nlweb.json
  Publish /.well-known/nlweb.json (Microsoft NLWeb) so agents can interact with the site through a natural-language interface.
  https://lucioduran.com/projects/ax-audit/guides/well-known-ai#nlweb

### Content Negotiation — 0/100 (reports, does not score)

- **FAIL** Homepage does not serve Markdown via content negotiation
  Got text/html (HTTP 200) for "Accept: text/markdown"
  Serve a Markdown representation of your pages when agents request "Accept: text/markdown". Agents like Claude Code and Cursor ask for it, and Markdown cuts token usage by ~80% vs HTML. Cloudflare ("Markdown for Agents") and Vercel can enable this without code changes.
  https://lucioduran.com/projects/ax-audit/guides/content-negotiation#not-supported
- **WARN** No <link rel="alternate" type="text/markdown"> fallback found on the homepage
  If you cannot enable content negotiation, advertise a Markdown version with <link rel="alternate" type="text/markdown" href="/index.md"> so agents can discover it.
  https://lucioduran.com/projects/ax-audit/guides/content-negotiation#no-alternate

### RSL License — 0/100 (reports, does not score)

- **FAIL** No RSL license discovery found
  Checked robots.txt License directive, Link header, and <link rel="license" type="application/rsl+xml">
  Declare machine-readable licensing terms for your content with Really Simple Licensing. Add to robots.txt: License: https://your-site.com/license.xml — then publish the RSL document. See https://rslstandard.org.
  https://lucioduran.com/projects/ax-audit/guides/rsl#not-found

### Agent Access — 100/100 (reports, does not score)

- **PASS** All 8 core AI crawler user-agents receive equivalent responses

### Structured Data — 80/100

- **PASS** 1 JSON-LD block(s) found
- **PASS** @context references schema.org
- **WARN** No @graph array (single-entity only)
  Use an @graph array to define multiple entities in one JSON-LD block: { "@context": "https://schema.org", "@graph": [...] }
  https://lucioduran.com/projects/ax-audit/guides/structured-data#no-graph
- **WARN** Only 1 key type found: WebPage
  Consider adding: Person, Organization, WebSite, ProfilePage
  Add more entity types to your @graph. AI agents use these to understand site structure. Common types: Person, Organization, WebSite, WebPage.
  https://lucioduran.com/projects/ax-audit/guides/structured-data#few-types
- **WARN** No BreadcrumbList found
  Add a BreadcrumbList entity to help AI agents understand your site navigation hierarchy.
  https://lucioduran.com/projects/ax-audit/guides/structured-data#no-breadcrumb

### HTTP Headers — 85/100

- **PASS** 4/7 security headers present
- **WARN** No Link header for AI discovery (llms.txt, agent.json)
  Add a Link response header pointing to your AI discovery files: Link: </llms.txt>; rel="alternate"; type="text/plain", </.well-known/agent.json>; rel="alternate"; type="application/json"
  https://lucioduran.com/projects/ax-audit/guides/http-headers#no-link-header

### HTML Rendering — 75/100

- **PASS** Server-rendered content detected (2599 words, 16181 chars of visible text)
- **WARN** Low text-to-markup ratio (2.5%)
  Recommended minimum: 5%
  A very low text-to-markup ratio is a typical SPA-shell symptom. Inline more content directly into the HTML response.
  https://lucioduran.com/projects/ax-audit/guides/html-rendering#low-ratio
- **PASS** Semantic landmarks present (main, article, section, header, footer, nav)
- **WARN** No <h1> heading found
  Add a single <h1> describing the page. Agents and search engines treat the H1 as the primary topic indicator.
  https://lucioduran.com/projects/ax-audit/guides/html-rendering#no-h1
- **WARN** Only 67/131 <img> tags have alt attributes
  Add descriptive alt="" to every <img>. Agents use alt text to understand images they cannot process visually.
  https://lucioduran.com/projects/ax-audit/guides/html-rendering#missing-alt

### SEO Basics — 90/100

- **WARN** <title> is too long (120 chars)
  "BBC Home - Breaking News, World News, US News, Sports, Business, Innovation, Cli…"
  Shorten the title to 20-70 characters. Many agents and search engines truncate beyond ~70.
  https://lucioduran.com/projects/ax-audit/guides/seo-basics#long-title
- **PASS** Meta description length 126 chars
- **WARN** 2 <link rel="canonical"> tags (must be exactly 1)
  Keep a single canonical link per page. Multiple canonical hints are ignored by agents and search engines.
  https://lucioduran.com/projects/ax-audit/guides/seo-basics#multiple-canonical
- **PASS** <html lang="en-GB">
- **PASS** UTF-8 charset declared
- **PASS** Viewport meta present: "width=device-width, initial-scale=1"

### Crawl Efficiency — 80/100 (reports, does not score)

- **PASS** Response compressed with gzip
  Brotli (br) typically compresses text 10–20% smaller — consider enabling it.
- **PASS** Cache validator present (ETag)
- **WARN** Conditional request returned 200 instead of 304 Not Modified
  Re-requested with If-None-Match
  The server advertises a cache validator but does not honor If-None-Match / If-Modified-Since. Configure it to return 304 when the validator matches, so crawlers avoid re-downloading unchanged pages.
  https://lucioduran.com/projects/ax-audit/guides/crawl-efficiency#no-304
- **WARN** Homepage is on the large side (630.8 KB decompressed)
  Consider trimming inlined payloads to reduce crawl cost.
  https://lucioduran.com/projects/ax-audit/guides/crawl-efficiency#large-page

---

Markdown representation of https://axrush.com/report/bbc.com
