Can AI assistants read dday.it?
Measured on 2026-09-19 by asking the site 7 times — once as a browser, once as each of the
6 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing
what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead.
B77 / 100
5 of 6 crawlers can read this page.
More readable than 57% of the 47,934 sites measured so far.
Structured data: NewsMediaOrganization.
What each crawler got back
| Crawler | HTTP | Result | robots.txt |
|---|
| ClaudeBot (Claude) | 200 |
can read it |
allowed-by-star |
| GPTBot (ChatGPT) | 200 |
allowed by the server, blocked in robots.txt |
blocked |
| OAI-SearchBot (ChatGPT Search) | 200 |
can read it |
allowed-by-star |
| PerplexityBot (Perplexity) | 200 |
can read it |
allowed-by-star |
| Google-Extended (Gemini, AI Overviews) a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own | — |
not opted out |
allowed-by-star |
| Meta-ExternalAgent (Meta AI) | 200 |
can read it |
allowed-by-star |
| Amazonbot (Alexa, Rufus) | 200 |
can read it |
allowed-by-star |
BunnyCDN answers for this domain. Where a crawler above is
refused while robots.txt allows it, the rule is applied by that edge, not written by the site — see
how often each edge does this.
What to change, in order
- Publish llms.txt
llms.txt is missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.
- Mark the content with <main>
The content-reachability check scores 2/100: with no <main> or <article> landmark, or with most of the page outside it, an assistant reads the navigation and the footer at the same cost as the text it came for.
- Fix the plain structure: one <h1>, a title, a description, alt text
The structure check scores 49/100. These are the cheapest signals on the page and the first ones an assistant uses to decide what the site is.
- 1 crawler is named and refused in robots.txt — a deliberate choice
GPTBot is blocked by a rule that names it. Nothing to fix here: this page records what is true, not what should be. It is listed so the deliberate part of the block is not confused with the accidental part above.
Get told if this changes
One email only when a measured crawler flips on dday.it, served to refused or back. No schedule, no newsletter; double opt-in, one-click stop.
The checks
- Answers AI agents like it answers people100%
- Readable without JavaScript100%
- robots.txt lets the crawlers in86%
- Facts in JSON-LD67%
- Content reachable, not buried2%
- Publishes a map of itself70%
- Plain structure49%
Full AI-visibility report for dday.it →
Scan your own site →
Sites with a similar score
Browse the whole index →
Embed this score
Put the badge on dday.it — it links back here, and re-measures every time this index re-crawls.
<a href="https://aivis.lumnika.com/ai-readiness/dday.it"><img src="https://aivis.lumnika.com/ai-readiness/dday.it/badge.svg" alt="AI readability"></a>