NeuralCrawl · August 19, 2026
NeuralCrawl
Monitor how companies and governments treat AI crawlers and search engine bots.
Analyze 1273 major US & European companies, governments, social networks and publishers (updated 24/7).
160 of 1196 analysed websites explicitly block at least one AI crawler by name in their robots.txt
GPTBot (OpenAI) — 111 companies
AI bots rejection over time weekly · US 500 cohort
12 weekly snapshots with reliable coverage
“13.4% of monitored organisations now explicitly block at least one AI crawler in their robots.txt
OpenAI's GPTBot is named by 111 of them.”
Most blocked AI crawlers share of 1196 analysed websites
Which sectors block AI the most all monitored sites · share blocking ≥1 AI crawler · min 5 companies
Most AI-restrictive companies AI crawlers blocked by name
| # | Company | Sector | Bots blocked |
|---|---|---|---|
| 4 | USA Today | News & Media | 28 |
| 60 | Toronto Star | News & Media | 28 |
| 15 | Politico | News & Media | 27 |
| 38 | Msn | Other | 27 |
| 44 | Daily Mail | News & Media | 27 |
| 104 | IGN | Other | 27 |
| 59 | The Globe and Mail | News & Media | 26 |
| 9 | CNN | News & Media | 25 |
| 42 | The Telegraph | News & Media | 25 |
| 56 | Nikkei | Finance & Insurance | 25 |
Change activity robots.txt modifications per week
ISO week numbers. Baseline (first-archive) events excluded.
Latest AI-policy moves from the change feed
- Reliance Steel & Aluminum — Blocked Bytespider (ByteDance) entirely
- Reliance Steel & Aluminum — Blocked CCBot (Common Crawl) entirely
- Reliance Steel & Aluminum — Blocked Claude-SearchBot (Anthropic) entirely
- Reliance Steel & Aluminum — Blocked ClaudeBot (Anthropic) entirely
- Reliance Steel & Aluminum — Blocked GPTBot (OpenAI) entirely
- Reliance Steel & Aluminum — Blocked meta-externalagent (Meta) entirely
- Reliance Steel & Aluminum — Blocked omgili (Webz.io) entirely
- Renault Group — Unblocked ClaudeBot (Anthropic)
- Renault Group — Removed Allow: /br/ for ClaudeBot
- Renault Group — Removed Allow: /de/ for ClaudeBot
- Renault Group — Removed Allow: /es/ for ClaudeBot
- Renault Group — Removed Allow: /it/ for ClaudeBot
- Renault Group — Removed Allow: /pt/ for ClaudeBot
- Renault Group — Removed Allow: /ro/ for ClaudeBot
- Renault Group — Unblocked GPTBot (OpenAI)
- Renault Group — Removed Allow: /br/ for GPTBot
- Renault Group — Removed Allow: /de/ for GPTBot
- Renault Group — Removed Allow: /es/ for GPTBot
- Renault Group — Removed Allow: /it/ for GPTBot
- Renault Group — Removed Allow: /pt/ for GPTBot
- Renault Group — Removed Allow: /ro/ for GPTBot
- DOAJ — Unblocked Amazonbot (Amazon) - rules removed
- DOAJ — Unblocked Applebot-Extended (Apple) - rules removed
- DOAJ — Unblocked Bytespider (ByteDance) - rules removed
- DOAJ — Unblocked CCBot (Common Crawl) - rules removed
- DOAJ — Unblocked ClaudeBot (Anthropic) - rules removed
- DOAJ — Unblocked Google-Extended (Google) - rules removed
- DOAJ — Unblocked GPTBot (OpenAI) - rules removed
- DOAJ — Unblocked meta-externalagent (Meta) - rules removed
- Indeed — Added Allow: /*&start=0& for AI2Bot
- Indeed — Added Allow: /*&start=10& for AI2Bot
- Indeed — Added Allow: /*&start=20& for AI2Bot
- Indeed — Added Allow: /*&start=30& for AI2Bot
- Indeed — Added Allow: /*&start=40& for AI2Bot
- Indeed — Added Allow: /*&start=50& for AI2Bot
- Indeed — Added Allow: /*&start=60& for AI2Bot
- Indeed — Added Allow: /*&start=70& for AI2Bot
- Indeed — Added Allow: /*&start=80& for AI2Bot
- Indeed — Added Allow: /*&start=90& for AI2Bot
- Indeed — Added Allow: /*&start=0& for AmazonBot
- Indeed — Added Allow: /*&start=10& for AmazonBot
- Indeed — Added Allow: /*&start=20& for AmazonBot
- Indeed — Added Allow: /*&start=30& for AmazonBot
- Indeed — Added Allow: /*&start=40& for AmazonBot
- Indeed — Added Allow: /*&start=50& for AmazonBot
- Indeed — Added Allow: /*&start=60& for AmazonBot
- Indeed — Added Allow: /*&start=70& for AmazonBot
- Indeed — Added Allow: /*&start=80& for AmazonBot
- Indeed — Added Allow: /*&start=90& for AmazonBot
- Indeed — Added Allow: /*&start=0& for anthropic-ai
- Indeed — Added Allow: /*&start=10& for anthropic-ai
- Indeed — Added Allow: /*&start=20& for anthropic-ai
- Indeed — Added Allow: /*&start=30& for anthropic-ai
- Indeed — Added Allow: /*&start=40& for anthropic-ai
- Indeed — Added Allow: /*&start=50& for anthropic-ai
- Indeed — Added Allow: /*&start=60& for anthropic-ai
- Indeed — Added Allow: /*&start=70& for anthropic-ai
- Indeed — Added Allow: /*&start=80& for anthropic-ai
- Indeed — Added Allow: /*&start=90& for anthropic-ai
- Indeed — Added Allow: /*&start=0& for Applebot-Extended
- Indeed — Added Allow: /*&start=10& for Applebot-Extended
- Indeed — Added Allow: /*&start=20& for Applebot-Extended
- Indeed — Added Allow: /*&start=30& for Applebot-Extended
- Indeed — Added Allow: /*&start=40& for Applebot-Extended
- Indeed — Added Allow: /*&start=50& for Applebot-Extended
- Indeed — Added Allow: /*&start=60& for Applebot-Extended
- Indeed — Added Allow: /*&start=70& for Applebot-Extended
- Indeed — Added Allow: /*&start=80& for Applebot-Extended
- Indeed — Added Allow: /*&start=90& for Applebot-Extended
- AstraZeneca — Added rules for AI crawler Amazonbot (Amazon)
- AstraZeneca — Added rules for AI crawler Applebot-Extended (Apple)
- AstraZeneca — Added rules for AI crawler CCBot (Common Crawl)
- AstraZeneca — Added rules for AI crawler ChatGPT-User (OpenAI)
- AstraZeneca — Added rules for AI crawler Claude-SearchBot (Anthropic)
- AstraZeneca — Added rules for AI crawler Claude-User (Anthropic)
- AstraZeneca — Added rules for AI crawler ClaudeBot (Anthropic)
- AstraZeneca — Added rules for AI crawler DuckAssistBot (DuckDuckGo)
- AstraZeneca — Added rules for AI crawler Google-Extended (Google)
- AstraZeneca — Added rules for AI crawler GPTBot (OpenAI)
- AstraZeneca — Added rules for AI crawler meta-externalagent (Meta)
- AstraZeneca — Added rules for AI crawler MistralAI-User (Mistral AI)
- AstraZeneca — Added rules for AI crawler OAI-SearchBot (OpenAI)
- AstraZeneca — Added rules for AI crawler Perplexity-User (Perplexity)
- AstraZeneca — Added rules for AI crawler PerplexityBot (Perplexity)
- AstraZeneca — Added rules for AI crawler YouBot (You.com)
- Itaú Unibanco — Added rules for AI crawler Amazonbot (Amazon)
- Itaú Unibanco — Added rules for AI crawler Applebot-Extended (Apple)
- Itaú Unibanco — Added rules for AI crawler CCBot (Common Crawl)
- Itaú Unibanco — Added rules for AI crawler Google-Extended (Google)
- Itaú Unibanco — Added rules for AI crawler meta-externalagent (Meta)
- Itaú Unibanco — Removed rules for AI crawler ChatGPT-User (OpenAI)
- Itaú Unibanco — Removed rules for AI crawler Claude-User (Anthropic)
- Itaú Unibanco — Removed rules for AI crawler Perplexity-User (Perplexity)
- Itaú Unibanco — Removed rules for AI crawler YouBot (You.com)
- Itaú Unibanco — Added Allow: /investimentos/ for ClaudeBot
- Itaú Unibanco — Added Disallow: /api/internal/ for ClaudeBot
- Itaú Unibanco — Added Disallow: /auth/ for ClaudeBot
- Itaú Unibanco — Added Allow: /investimentos/ for OAI-SearchBot
- Itaú Unibanco — Added Disallow: /api/internal/ for OAI-SearchBot
- Itaú Unibanco — Added Disallow: /login/ for OAI-SearchBot
- Itaú Unibanco — Added Disallow: /private/ for OAI-SearchBot
- Itaú Unibanco — Added Allow: /investimentos/ for PerplexityBot
- Itaú Unibanco — Added Disallow: /api/internal/ for PerplexityBot
- Itaú Unibanco — Added Disallow: /auth/ for PerplexityBot
- Peru (Gob.pe) — Added rules for AI crawler GPTBot (OpenAI)
- Peru (Gob.pe) — Added rules for AI crawler OAI-SearchBot (OpenAI)
- Figma Community — Added Allow: /blog/how-to-move-fast-toward-the-right-thing/$ for CHATGPT-User
- Figma Community — Added Allow: /blog/how-to-move-fast-toward-the-right-thing/$ for Claude-SearchBot
- Figma Community — Added Allow: /blog/how-to-move-fast-toward-the-right-thing/$ for Claude-User
- Figma Community — Added Allow: /blog/how-to-move-fast-toward-the-right-thing/$ for OAI-SearchBot
- Figma Community — Added Allow: /blog/how-to-move-fast-toward-the-right-thing/$ for PerplexityBot
Methodology. robots.txt files are fetched every 6 hours from the
primary domains of every monitored cohort (US & European companies,
governments and social networks). "Blocks" counts organisations that name an
AI crawler in their own User-agent group with
Disallow: /. 1196 of 1273 sites are reachable
and analysed. Browse cohorts on the datasets page and
trends charts. Full crawl health on the
status page.