Agent Web Index › lonelyplanet.com

Can AI assistants read lonelyplanet.com?

Measured on 2026-09-20 by asking the site 9 times — once as a browser, once as each of the 8 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead. Tranco rank #3,694.

B78 / 100
6 of 8 crawlers can read this page.
More readable than 60% of the 47,934 sites measured so far. Structured data: Organization, WebSite, WebPage.

What each crawler got back

CrawlerHTTPResultrobots.txt
ClaudeBot (Claude)200 can read it allowed-by-star
GPTBot (ChatGPT)403 challenged blocked
OAI-SearchBot (ChatGPT Search)200 can read it allowed-by-star
PerplexityBot (Perplexity)200 can read it allowed-by-star
Google-Extended (Gemini, AI Overviews)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed-by-star
Meta-ExternalAgent (Meta AI)200 can read it allowed-by-star
Amazonbot (Alexa, Rufus)200 can read it allowed-by-star
Bytespider (Doubao, Lark)403 challenged allowed-by-star
Applebot (Siri, Apple Intelligence)200 can read it allowed-by-star
Applebot-Extended (Apple Intelligence training)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed-by-star

AWS CloudFront answers for this domain. Where a crawler above is refused while robots.txt allows it, the rule is applied by that edge, not written by the site — see how often each edge does this.

What to change, in order

  1. Let in the 1 crawler your own robots.txt already allows+1 crawler
    Bytespider is turned away before reading the page (served a bot challenge instead of the page), while robots.txt permits it — so this block is not written in your site. AWS CloudFront answers for this domain. AWS WAF Bot Control has an AI-crawler category that CloudFront enforces ahead of your origin. https://docs.aws.amazon.com/waf/latest/developerguide/waf-bot-control.html Fixing it takes this domain from 6 to 7 of 8 crawlers.
  2. Publish llms.txt
    llms.txt is missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.
  3. Mark the content with <main>
    The content-reachability check scores 16/100: with no <main> or <article> landmark, or with most of the page outside it, an assistant reads the navigation and the footer at the same cost as the text it came for.
  4. 1 crawler is named and refused in robots.txt — a deliberate choice
    GPTBot is blocked by a rule that names it. Nothing to fix here: this page records what is true, not what should be. It is listed so the deliberate part of the block is not confused with the accidental part above.

Get told if this changes

One email only when a measured crawler flips on lonelyplanet.com, served to refused or back. No schedule, no newsletter; double opt-in, one-click stop.

The checks

History

2026-09-13: 78 · 2026-09-14: 78 · 2026-09-15: 78 · 2026-09-16: 78 · 2026-09-17: 78 · 2026-09-18: 78 · 2026-09-19: 80 · 2026-09-20: 78

Full AI-visibility report for lonelyplanet.com → Scan your own site →

Sites with a similar score

kibernet.huB 79locknlock.inB 78looknbookart.comB 78liu.seB 78lqibev.comB 78lelscans.netB 78maruzenjunkudo.co.jpB 78katanweaves.comB 78mozello.comB 78ezeller.comB 78restream.ioB 78ogolosha.uaB 77

Browse the whole index →

Embed this score

Put the badge on lonelyplanet.com — it links back here, and re-measures every time this index re-crawls.

lonelyplanet.com AI readability: B 78/100
<a href="https://aivis.lumnika.com/ai-readiness/lonelyplanet.com"><img src="https://aivis.lumnika.com/ai-readiness/lonelyplanet.com/badge.svg" alt="AI readability"></a>
Method. 9 live HTTP requests (one per crawler, one as a browser) plus robots.txt, llms.txt, sitemap.xml and security.txt, 12-second timeout each, from a single vantage point. Blocking AI crawlers is a legitimate choice, not a failure: this page records what is true, not what should be. Domain comes from the Tranco research list. One page per domain — the homepage — is audited.

Part of the Agent Web Index, 47,934 domains measured, updated continuously.