NeuralCrawl

NeuralCrawl · September 15, 2026

NeuralCrawl

Monitor how companies and governments treat AI crawlers and search engine bots.

Analyze 1273 major US & European companies, governments, social networks and publishers (updated 24/7).

13.7% block AI bots

165 of 1201 analysed websites explicitly block at least one AI crawler by name in their robots.txt

9.6% block GPTBot (OpenAI) — 115 companies
3820 robots.txt snapshots in the archive
251 policy changes detected in the last 7 days
1204 robots.txt files under 24/7 monitoring

AI bots rejection over time weekly · US 500 cohort

Block ≥1 AI crawler Block GPTBot
2026-09-04 2026-09-15

12 weekly snapshots with reliable coverage

“13.7% of monitored organisations now explicitly block at least one AI crawler in their robots.txt
OpenAI's GPTBot is named by 115 of them.”

Most blocked AI crawlers share of 1201 analysed websites

Which sectors block AI the most all monitored sites · share blocking ≥1 AI crawler · min 5 companies

News & Media 56.4% (44/78)
Social & Communities 50.0% (6/12)
Entertainment & Sports 26.5% (9/34)
Telecom 16.7% (4/24)
Retail & E-commerce 12.7% (13/102)
Education & Research 10.4% (7/67)
Technology & Software 10.1% (24/238)
Energy & Utilities 7.0% (5/71)
Travel & Hospitality 4.8% (1/21)
Finance & Insurance 4.3% (6/139)

Most AI-restrictive companies AI crawlers blocked by name

#CompanySectorBots blocked
4 USA Today News & Media 28
60 Toronto Star News & Media 28
15 Politico News & Media 27
20 DBLP Other 27
38 Msn Other 27
44 Daily Mail News & Media 27
104 IGN Other 27
59 The Globe and Mail News & Media 26
9 CNN News & Media 25
42 The Telegraph News & Media 25

Change activity robots.txt modifications per week

202 26
201 27
198 28
238 29
234 30
221 31
219 32
215 33
243 34
201 35
224 36
54 37

ISO week numbers. Baseline (first-archive) events excluded.

Latest AI-policy moves from the change feed

Methodology. robots.txt files are fetched every 6 hours from the primary domains of every monitored cohort (US & European companies, governments and social networks). "Blocks" counts organisations that name an AI crawler in their own User-agent group with Disallow: /. 1201 of 1273 sites are reachable and analysed. Browse cohorts on the datasets page and trends charts. Full crawl health on the status page.