Agent Web Index › cian.ru

Can AI assistants read cian.ru?

Measured on 2026-09-20 by asking the site 9 times — once as a browser, once as each of the 8 crawlers that feed ChatGPT, Claude, Perplexity, Gemini, Meta AI, Apple Intelligence and Doubao — and comparing what came back. The AI-training opt-out tokens that never crawl are read from robots.txt instead. Tranco rank #3,485.

C66 / 100
2 of 8 crawlers can read this page.
More readable than 27% of the 47,934 sites measured so far. Structured data: LodgingBusiness.

What each crawler got back

CrawlerHTTPResultrobots.txt
ClaudeBot (Claude)403 challenged allowed-by-star
GPTBot (ChatGPT)403 challenged allowed-by-star
OAI-SearchBot (ChatGPT Search)403 challenged allowed-by-star
PerplexityBot (Perplexity) no-answer allowed-by-star
Google-Extended (Gemini, AI Overviews)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed-by-star
Meta-ExternalAgent (Meta AI)200 can read it allowed-by-star
Amazonbot (Alexa, Rufus)200 thin allowed-by-star
Bytespider (Doubao, Lark)403 challenged allowed-by-star
Applebot (Siri, Apple Intelligence)200 can read it allowed-by-star
Applebot-Extended (Apple Intelligence training)
a robots.txt token, not a crawler: it controls how already-crawled pages may be used, and never makes a request of its own
not opted out allowed-by-star

What to change, in order

  1. Let in the 6 crawlers your own robots.txt already allows+6 crawlers
    ClaudeBot, GPTBot, OAI-SearchBot, PerplexityBot, Amazonbot, Bytespider are turned away before reading the page (served a bot challenge instead of the page; the connection never completed; answered, but with far less text than a browser gets), while robots.txt permits them — so this block is not written in your site. No CDN signature was found in the response headers, so the refusal comes from the origin server itself or from a WAF this index does not recognise. Fixing it takes this domain from 2 to 8 of 8 crawlers.
  2. Publish llms.txt
    llms.txt is missing, so an assistant has to discover the site by following links. llms.txt is the emerging convention for telling an assistant which pages actually matter.

Get told if this changes

One email only when a measured crawler flips on cian.ru, served to refused or back. No schedule, no newsletter; double opt-in, one-click stop.

The checks

History

2026-09-13: 40 · 2026-09-14: 40 · 2026-09-15: 40 · 2026-09-16: 40 · 2026-09-17: 40 · 2026-09-18: 59 · 2026-09-19: 63 · 2026-09-20: 66

Full AI-visibility report for cian.ru → Scan your own site →

Sites with a similar score

creaders.netC 67chura.myshopify.comC 66cigars.myshopify.comC 66chineseherbalpharmacy.myshopify.comC 66cleanair.myshopify.comC 66cdmwebs.myshopify.comC 66communicatie.myshopify.comC 66blueiguana.myshopify.comC 66desert-blossom.myshopify.comC 66hp31.myshopify.comC 66wickedquiver.myshopify.comC 65kshop.myshopify.comC 65

Browse the whole index →

Embed this score

Put the badge on cian.ru — it links back here, and re-measures every time this index re-crawls.

cian.ru AI readability: C 66/100
<a href="https://aivis.lumnika.com/ai-readiness/cian.ru"><img src="https://aivis.lumnika.com/ai-readiness/cian.ru/badge.svg" alt="AI readability"></a>
Method. 9 live HTTP requests (one per crawler, one as a browser) plus robots.txt, llms.txt, sitemap.xml and security.txt, 12-second timeout each, from a single vantage point. Blocking AI crawlers is a legitimate choice, not a failure: this page records what is true, not what should be. Domain comes from the Tranco research list. One page per domain — the homepage — is audited.

Part of the Agent Web Index, 47,934 domains measured, updated continuously.