Skip to content
Home » Do News Websites Block AI Crawlers? I Tested 13 Major Publishers

Do News Websites Block AI Crawlers? I Tested 13 Major Publishers

Original Research: This AI Visibility Benchmark evaluated thirteen leading news and media websites using the same testing method on every article. One tool bug turned up mid-testing, and I caught it before it reached this page. The numbers below are the corrected results.

Everyone assumes major news outlets are ready for AI search. Nobody actually checks.

I picked thirteen well-known news and media publishers, one AI-related article from each, and ran every single one through the same tools, on the same day, using the same method. No guessing. No borrowed data. Just the numbers.

What I found broke almost every assumption I walked in with.

Research Summary

Thirteen news and media publishers tested. Twenty-six AI Visibility and Citability measurements, plus a full seven-bot crawler check on every site. Same method for every article. One tool bug found, caught, and worked around mid-testing.

Benchmark at a Glance

Highest AI Visibility score: 75 (Wired). Highest Citability score: 86 (Wired). Most open to AI crawlers: Fast Company and Inc, zero of seven bots blocked. Most restricted: BBC, all seven bots blocked.

Average AI Visibility: 50. Average Citability: 48. Average crawler blocks: 3.8 of seven bots per site, about 54 percent of the crawler set.

Here is the number that actually surprised me. The two sites that first showed a perfect crawler score turned out to be wrong. I will get to that.

Key Findings

Nine of the thirteen publishers I tested block at least four of the seven AI crawlers I checked. Wired led the whole benchmark on both AI Visibility and Citability, the only site to top both charts.

BBC produced the third-most citable content in the set, 69 out of 100, while blocking every single AI crawler I tested. Fast Company and Inc, both owned by Mansueto Ventures, share a robots.txt file with the exact same outdated bot names, meaning most of their supposed AI blocks do nothing at all.

Reuters, Fast Company, and Inc all scored 6 on Citability, the weakest content-quality results in the set. My own AI robots.txt Checker gave two sites a false perfect score, and I only caught it by checking the raw files by hand.

Why I Built This Benchmark

Every major outlet now covers AI constantly. Fewer of them have checked whether AI can actually read what they publish.

I wanted real numbers, not assumptions. Not “big publishers must have this handled.” Just what showed up when I tested the actual pages.

Benchmark Methodology

Why These Thirteen Websites

I picked thirteen sites that represent a real spread of the news and media industry: general news, business news, and tech-focused outlets. All thirteen have real audiences and real editorial standards. All thirteen have published AI-related coverage recently, which gave me a like-for-like article to test on each one.

Articles Tested

WebsiteArticle TestedTopic
TechCrunchGoogle’s AI Search Is Rapidly Becoming the DefaultAI search adoption
ForbesPredicting AI In 2026: A Year Of ConsequenceAI industry outlook
Fast CompanyCES 2026: The Year AI Got SeriousAI trends
Inc10 Small Business Ideas With Six-Figure PotentialAI and small business
CNBCNvidia and OpenAI in Talks for Up to $250 Billion AI BackstopAI infrastructure
The GuardianGrok Image Generator Turned Off After OutcryAI safety
Ars TechnicaPerplexity Announces Computer, an AI Agent That Assigns Work to Other AI AgentsAI agents
ReutersOpenAI to Nearly Double Workforce to 8,000 by End of 2026AI industry growth
Associated PressWhite House Urges Congress to Take a Light Touch on AI RegulationAI policy
Business InsiderExecutives Share 2026 AI PredictionsAI industry outlook
BBCAI Chatbots Unable to Accurately Summarise News, BBC FindsAI and news accuracy
WiredThe OpenAI and Anthropic AI Hacking Sprees Are a Messy New Legal FrontierAI and law
The VergeGoogle Earth’s AI Deepfake Tool Only Lasted One DayAI safety

Testing Process

I ran every article through the same tools, in the same order, on the same day. One article per site, no retesting to chase a better number, and no manual scoring.

Testing date: August 2, 2026. All thirteen articles were tested on the same day. Crawler-access findings reflect the robots.txt files available at the time of testing.

The Tools I Used, and the Bug I Found Mid-Testing

AI Visibility Checker

Measures overall AI readiness: technical signals, structural elements, and metadata quality. Returns a score out of 100.

Content Citability Grader

Measures how likely a piece of content is to get cited by an AI system, across Evidence, Structure, Authority, and AI Readability. Each dimension is worth 25 points.

AI Crawler Access, Checked By Hand

I originally planned to run my own AI robots.txt Checker as the third tool, same as I did with Schema on my last benchmark. Partway through testing, I found it misreports certain robots.txt structures, showing blocked bots as allowed. Two sites, The Guardian and Ars Technica, both showed false 100 out of 100 scores before I caught it.

I confirmed the bug using an independent robots.txt parser, then manually checked all seven bots, GPTBot, ClaudeBot, PerplexityBot, Google-Extended, OAI-SearchBot, CCBot, and anthropic-ai, on every one of the thirteen sites. The crawler numbers in this piece come from that manual check, not my own tool.

That is worth sitting with for a second. My own site caught its own tool getting this wrong. I would rather tell you that than quietly fix it and pretend it never happened.

Overall Benchmark Results

WebsiteAI VisibilityContent CitabilityBots Blocked (of 7)
Wired75864
The Guardian69534
Ars Technica65545
AP59525
Business Insider59593
The Verge59665
CNBC48496
Forbes45701
BBC45697
Reuters4464
TechCrunch32495
Fast Company2460
Inc2460
Average50483.8

No two sites showed the same profile. Some led on content. Some led on access. Almost none led on both.

AI Visibility Hall of Fame

Best Overall AI Visibility, Wired, 75. The highest raw score in the whole benchmark, and it held up across every signal I checked.

Best Content Citability, Wired, 86, rated Excellent by my own grader. No other site cleared 70.

Most Open to AI Crawlers, Fast Company and Inc, zero of seven bots blocked. Worth a caveat here: most of those blocks were written against outdated bot names that no longer match real crawlers, so the openness is partly an accident, not a deliberate choice.

Most Restricted, BBC, seven of seven bots blocked, the only fully closed site in the set.

Biggest Surprise, BBC again. Its content scored 69 on Citability, third-best in the whole benchmark, while its crawlers score sits at zero access. One of the most citable articles I tested is also the least reachable one.

Most Interesting Outlier, Fast Company and Inc, sharing one broken robots.txt template between two properties owned by the same parent company, Mansueto Ventures.

Overall AI Visibility Scores

AI Visibility Checker results, thirteen news and media sites, same method, same day.

RankWebsiteAI Visibility Score
1Wired75
2The Guardian69
3Ars Technica65
4AP59
4Business Insider59
4The Verge59
7CNBC48
8Forbes45
8BBC45
10Reuters44
11TechCrunch32
12Fast Company24
13Inc24
Benchmark Average50

Tested using the Free AI Visibility Checker at nenawow.com, same method, same day, one article per site.

Content Citability Scores

Content Citability Grader results, and this is where the real spread showed up.

RankWebsiteContent Citability Score
1Wired86
2Forbes70
3BBC69
4The Verge66
5Business Insider59
6Ars Technica54
7The Guardian53
8AP52
9TechCrunch49
9CNBC49
11Fast Company6
12Inc6
13Reuters6
Benchmark Average48

TechCrunch and CNBC landed close together, both near 49, while Fast Company, Inc, and Reuters all landed at exactly 6. That is not noise. That is a real pattern worth a closer look.

Individual Website Analysis

TechCrunch

AI Visibility 32, Citability 49, five of seven bots blocked. TechCrunch writes about AI constantly, including the article I tested here, and still comes back Heavily Restricted on crawler access.

NenaWow's Free AI Visibility Checker results for TechCrunch's AI search adoption article showing an Overall AI Visibility Score of 32 out of 100, rated Poor
TechCrunch’s article on AI search adoption earned an AI Visibility Score of just 32/100 in the benchmark, rated Poor, well below the 68 average from the SEO-sites benchmark.
NenaWow's Free Content Citability Grader results for TechCrunch's AI search adoption article showing an Overall Citability Score of 49 out of 100, rated Poor
TechCrunch’s AI search adoption article scored 49/100 on Content Citability, rated Poor, well below the 78 average from the SEO-sites benchmark.

TechCrunch blocks GPTBot, ClaudeBot, Google-Extended, CCBot, and anthropic-ai, while letting Perplexity through. That is a deliberate, specific choice, not a technical accident. A publisher covering AI search as a beat is also one of the harder sites in this set for AI systems to actually reach.

Research conclusion: TechCrunch’s coverage of AI outpaces its own openness to it.

Forbes

AI Visibility 45, Citability 70, just one bot blocked. Forbes produced the second-highest Citability score in the set at 70, one point ahead of BBC.

NenaWow's Free AI Visibility Checker results for Forbes' AI in 2026 outlook article showing an Overall AI Visibility Score of 45 out of 100, rated Fair
Forbes’ 2026 AI outlook article earned an AI Visibility Score of 45/100, rated Fair, ahead of TechCrunch’s 32 but still below the 68 average from the SEO-sites benchmark.
NenaWow's Free Content Citability Grader results for Forbes' AI in 2026 outlook article showing an Overall Citability Score of 70 out of 100, rated Good
Forbes’ AI 2026 outlook article scored 70/100 on Content Citability, rated Good, close to the 78 average from the SEO-sites benchmark and well ahead of TechCrunch’s 49.

That combination, strong content and near-full crawler access, makes Forbes one of the more balanced performers here. It never leads a single category, but it never falls near the bottom either.

Research conclusion: Forbes is the steadiest, least dramatic result in the whole benchmark.

Fast Company

AI Visibility 24, Citability 6, zero bots blocked, technically. Fast Company’s raw robots.txt file lists eleven AI-related blocks, but most target crawler names that no longer match how these bots actually identify themselves.

NenaWow's Free AI Visibility Checker results for Fast Company's CES 2026 AI trends article showing an Overall AI Visibility Score of 24 out of 100, rated Poor
Fast Company’s CES 2026 AI trends article earned an AI Visibility Score of 24/100, rated Poor, tying Inc for the lowest AI Visibility score in this benchmark.
NenaWow's Free Content Citability Grader results for Fast Company's CES 2026 AI trends article showing an Overall Citability Score of 6 out of 100, rated Very Low
Fast Company’s CES 2026 AI trends article scored just 6/100 on Content Citability, rated Very Low, the weakest citability result in either benchmark so far, far below Search Engine Journal’s previous low of 59.

Only its content scores are real weak points here. A 6 on Citability is close to the floor of the whole scale, and 24 on AI Visibility puts it among the lowest in the set.

Research conclusion: Fast Company’s openness may be accidental, since the shared rules target crawler names that no longer match real bots. Its content still needs real rebuilding.

Inc

AI Visibility 24, Citability 6, zero bots blocked, same as Fast Company. Inc shares its parent company, Mansueto Ventures, with Fast Company, and it shares something else too: the identical outdated robots.txt template, right down to the same non-functional bot names.

NenaWow's Free AI Visibility Checker results for Inc's small business ideas article showing an Overall AI Visibility Score of 24 out of 100, rated Poor
Inc’s small business ideas article earned an AI Visibility Score of 24/100, rated Poor, tying Fast Company as the weakest AI Visibility result in the benchmark.
NenaWow's Free Content Citability Grader results for Inc's small business ideas article showing an Overall Citability Score of 6 out of 100, rated Very Low
Inc’s small business ideas article scored 6/100 on Content Citability, rated Very Low, matching Fast Company’s score as the lowest citability result in the benchmark.

Two different newsrooms, two different articles, and one shared technical file with the same outdated bot names. That is a rare, specific finding, not a coincidence.

Research conclusion: Inc’s numbers are Fast Company’s numbers, because the underlying setup is the same file.

CNBC

AI Visibility 48, Citability 49, six of seven bots blocked. CNBC’s file is clean and correctly written, using the exact current names for every major crawler, and it blocks nearly all of them anyway.

NenaWow's Free AI Visibility Checker results for CNBC's Nvidia and OpenAI AI backstop article showing an Overall AI Visibility Score of 48 out of 100, rated Fair
CNBC’s Nvidia and OpenAI article earned an AI Visibility Score of 48/100, rated Fair, slightly above Forbes and BBC at 45.
NenaWow's Free Content Citability Grader results for CNBC's Nvidia and OpenAI article showing an Overall Citability Score of 49 out of 100, rated Poor
CNBC’s Nvidia and OpenAI article scored 49/100 on Content Citability, rated Poor, matching TechCrunch’s score exactly.

That is the sharpest contrast in this section of the benchmark. A financial outlet reporting on a 250 billion dollar AI infrastructure story, fully and correctly walling itself off from the AI systems that story is about.

Research conclusion: CNBC’s block is real, deliberate, and almost complete.

The Guardian

AI Visibility 69, Citability 53, four of seven bots blocked, once I corrected my own tool’s false reading. The Guardian’s robots.txt technically allows Google-Extended, but genuinely blocks ClaudeBot and PerplexityBot, not the perfect access my checker first reported.

NenaWow's Free AI Visibility Checker results for The Guardian's Grok image generator article showing an Overall AI Visibility Score of 69 out of 100, rated Good
The Guardian’s Grok image generator article earned an AI Visibility Score of 69/100, rated Good, the second-highest AI Visibility result in the benchmark behind Wired.
NenaWow's Free Content Citability Grader results for The Guardian's Grok image generator article showing an Overall Citability Score of 53 out of 100, rated Fair
The Guardian’s Grok image generator article scored 53/100 on Content Citability, rated Fair, its second-best result behind AI Visibility.

Even more worth reading closely: the file opens with a plain-English notice stating that AI and LLM use of Guardian content is not permitted without a license. Being crawlable and being authorized are two different things, and this is the clearest example of that gap I found anywhere in the set.

Research conclusion: The Guardian is technically reachable and legally closed at the same time.

Ars Technica

AI Visibility 65, Citability 54, five of seven bots blocked, once corrected. Same story as The Guardian: my own tool first reported a perfect 100, and a manual check found ClaudeBot, PerplexityBot, and Google-Extended all genuinely blocked.

NenaWow's Free AI Visibility Checker results for Ars Technica's Perplexity Computer article showing an Overall AI Visibility Score of 65 out of 100, rated Good
Ars Technica’s Perplexity Computer article earned an AI Visibility Score of 65/100, rated Good, close behind The Guardian’s 69.
NenaWow's Free Content Citability Grader results for Ars Technica's Perplexity Computer article showing an Overall Citability Score of 54 out of 100, rated Fair
Ars Technica’s Perplexity Computer article scored 54/100 on Content Citability, rated Fair, nearly identical to The Guardian’s 53.

Ars Technica’s raw file stacks a long list of user-agents under one shared Disallow rule, a pattern my checker apparently cannot parse correctly. Worth knowing if you run your own robots.txt audit and see a similar structure.

Research conclusion: Ars Technica is more restricted than any automated first read would tell you.

Reuters

AI Visibility 44, Citability 6, four of seven bots blocked. Reuters ties Fast Company and Inc at the very bottom of the Citability scale, a real, specific weak point given how much AI-industry news Reuters actually breaks.

NenaWow's Free AI Visibility Checker results for Reuters' OpenAI workforce article showing an Overall AI Visibility Score of 44 out of 100, rated Fair
Reuters’ OpenAI workforce expansion article earned an AI Visibility Score of 44/100, rated Fair, close to CNBC’s 48 and Forbes’ 45.
NenaWow's Free Content Citability Grader results for Reuters' OpenAI workforce article showing an Overall Citability Score of 6 out of 100, rated Very Low
Reuters’ OpenAI workforce article scored 6/100 on Content Citability, rated Very Low, tying Fast Company and Inc as the weakest citability results in the benchmark.

That gap is worth naming plainly, though with just one Reuters article tested, this reads as a pattern in this sample, not a verdict on wire journalism as a whole.

Research conclusion: this Reuters article was fast. It was not citable.

Associated Press

AI Visibility 59, Citability 52, five of seven bots blocked. AP breaks the low-citability pattern that Reuters, Fast Company, and Inc all showed, landing in the middle of the pack instead of the floor.

NenaWow's Free AI Visibility Checker results for the Associated Press's White House AI regulation article showing an Overall AI Visibility Score of 59 out of 100, rated Fair
The Associated Press’s White House AI policy article earned an AI Visibility Score of 59/100, rated Fair, its third-highest AI Visibility result after The Guardian and Ars Technica.
NenaWow's Free Content Citability Grader results for the Associated Press's White House AI regulation article showing an Overall Citability Score of 52 out of 100, rated Fair
The Associated Press’s White House AI policy article scored 52/100 on Content Citability, rated Fair, close to CNBC’s 49.

So is straight news reporting always weak on citability? Not always. AP’s policy coverage carries more structure and sourcing than the other wire-style pieces I tested.

Research conclusion: AP proves the wire-format weakness is not universal.

Business Insider

AI Visibility 59, Citability 59, three of seven bots blocked. Business Insider is the only site in the whole benchmark where both scores landed on the exact same number.

NenaWow's Free AI Visibility Checker results for Business Insider's executive AI predictions article showing an Overall AI Visibility Score of 59 out of 100, rated Fair
Business Insider’s 2026 AI predictions article earned an AI Visibility Score of 59/100, rated Fair, matching AP exactly.
NenaWow's Free Content Citability Grader results for Business Insider's executive AI predictions article showing an Overall Citability Score of 59 out of 100, rated Fair
Business Insider’s 2026 AI predictions article scored 59/100 on Content Citability, matching its own AI Visibility score exactly at 59.

That symmetry is a small detail, not a major finding, but it is a clean one. Nothing here is a standout, and nothing here is a real weak point either.

Research conclusion: Business Insider is the definition of an average result.

BBC

AI Visibility 45, Citability 69, seven of seven bots blocked. BBC is the sharpest contrast in the entire benchmark. Its content sits second only to Wired on Citability, and it is the only site here blocking every major AI crawler completely.

NenaWow's Free AI Visibility Checker results for BBC's AI chatbot news accuracy article showing an Overall AI Visibility Score of 45 out of 100, rated Fair
BBC’s AI chatbot news accuracy article earned an AI Visibility Score of 45/100, rated Fair, tying Forbes exactly.
NenaWow's Free Content Citability Grader results for BBC's AI chatbot news accuracy article showing an Overall Citability Score of 69 out of 100, rated Fair
BBC’s AI chatbot news accuracy article scored 69/100 on Content Citability, the third-highest result in the benchmark, behind Wired at 86 and Forbes at 70.

One of the most citable articles I tested this cycle is also the one AI crawlers have the least chance of reaching. That is worth sitting with.

Research conclusion: BBC proves that great content and AI access are not the same problem.

Wired

AI Visibility 75, Citability 86, four of seven bots blocked. Wired leads the whole benchmark on both major scores, six points clear of The Guardian on Visibility and sixteen points clear of Forbes on Citability.

NenaWow's Free AI Visibility Checker results for Wired's AI hacking legal frontier article showing an Overall AI Visibility Score of 75 out of 100, rated Good
Wired’s article on the OpenAI and Anthropic AI hacking legal frontier earned an AI Visibility Score of 75/100, rated Good, the highest AI Visibility result in the entire benchmark.
NenaWow's Free Content Citability Grader results for Wired's AI hacking legal frontier article showing an Overall Citability Score of 86 out of 100, rated Excellent
Wired’s article on the OpenAI and Anthropic AI hacking legal frontier scored 86/100 on Content Citability, rated Excellent, the highest citability result in the entire benchmark by a wide margin.

Something specific about this article’s structure, its named legal sources, its direct definitions of a messy legal question, is driving that gap. This is the strongest single result in the whole set.

Research conclusion: Wired is the benchmark’s clear leader, not a close one.

The Verge

AI Visibility 59, Citability 66, five of seven bots blocked. The Verge lands mid-pack on Visibility but pulls ahead on Citability, its 66 sitting just behind BBC and Forbes.

NenaWow's Free AI Visibility Checker results for The Verge's Google Earth AI deepfake tool article showing an Overall AI Visibility Score of 59 out of 100, rated Fair
The Verge’s article on Google Earth’s AI deepfake tool earned an AI Visibility Score of 59/100, rated Fair, matching AP and Business Insider exactly.
NenaWow's Free Content Citability Grader results for The Verge's Google Earth AI deepfake tool article showing an Overall Citability Score of 66 out of 100, rated Fair
The Verge’s article on Google Earth’s AI deepfake tool scored 66/100 on Content Citability, rated Fair, its fourth-highest citability result in the benchmark.

That gap between the two scores is worth noting. Strong writing does not automatically pull crawler access or technical signals up with it.

Research conclusion: The Verge writes better than its technical setup gives it credit for.

Patterns I Discovered

These are the patterns that held up across thirteen pages, not conclusions about any one newsroom.

Content quality and crawler access measure completely different things

BBC scored 69 on Citability and got fully blocked. Fast Company scored 6 on Citability and left every crawler open. The two numbers do not move together, and treating them as one score would hide the real story on both sites.

A perfect crawler score deserves a second look, not blind trust

Two of the cleanest-looking results in this whole benchmark turned out to be wrong. If your own checker ever hands you a flat 100, read the raw file before you believe it.

Straight news writing varied widely on citability in this sample

Reuters, Fast Company, and Inc all landed at exactly 6. AP and Business Insider, covering similar ground with more structure and sourcing, both cleared 50. Worth watching whether this holds on a larger set, but format looks like it matters as much as topic here.

Shared ownership can mean shared mistakes

Fast Company and Inc post the same broken robots.txt file. If you run more than one property, check whether your technical setup is actually shared on purpose.

No two sites showed the same weakness

Some lead on content and lag on access. Some lead on access and lag on content. Nobody here got everything right.

Do News Websites Block AI Crawlers?

Manually verified results across all seven tracked bots, corrected after finding a bug in my own checker.

WebsiteGPTBotClaudeBotPerplexityBotGoogle-ExtendedOAI-SearchBotCCBotanthropic-ai
TechCrunchDisallowedDisallowedAllowedDisallowedAllowedDisallowedDisallowed
ForbesAllowedAllowedAllowedAllowedAllowedAllowedDisallowed
Fast CompanyAllowedAllowedAllowedAllowedAllowedAllowedAllowed
IncAllowedAllowedAllowedAllowedAllowedAllowedAllowed
CNBCDisallowedDisallowedDisallowedAllowedDisallowedDisallowedDisallowed
The GuardianAllowedDisallowedDisallowedAllowedAllowedDisallowedDisallowed
ReutersAllowedDisallowedDisallowedAllowedAllowedDisallowedDisallowed
Associated PressDisallowedDisallowedDisallowedAllowedAllowedDisallowedDisallowed
Business InsiderAllowedDisallowedAllowedAllowedAllowedDisallowedDisallowed
Ars TechnicaAllowedDisallowedDisallowedDisallowedAllowedDisallowedDisallowed
BBCDisallowedDisallowedDisallowedDisallowedDisallowedDisallowedDisallowed
WiredAllowedDisallowedDisallowedDisallowedAllowedDisallowedAllowed
The VergeAllowedDisallowedDisallowedDisallowedAllowedDisallowedDisallowed

Nine of the thirteen publishers blocked at least four of the seven AI crawlers I tracked. Across the full benchmark, the average publisher blocked 3.8 of seven, about 54 percent of the crawler set. That is not a fringe pattern. Nine of the thirteen publishers in this sample blocked a majority of the AI crawlers I tested.

What Website Owners Can Learn

The lessons here apply whether you run a major newsroom or a single review site.

Check your robots.txt by hand, not just through one tool. Two of my own results were wrong until I verified them independently. If a score looks too clean, it probably is.

Great content and open crawler access are two separate jobs. BBC nailed one and missed the other completely. Do not assume fixing one fixes both.

Shared infrastructure needs its own audit. If you run more than one site under one company, check whether your technical files are actually identical, and whether that was ever a real decision.

Straight reporting can still be citable. AP’s result suggests the difference comes down to sourcing and structure more than the news format itself, though this is one article, not a formula.

Limitations

I found a bug in my own AI robots.txt Checker mid-testing. It misreports certain multi-agent block structures in a robots.txt file, showing genuinely blocked bots as allowed. I confirmed this with an independent parser and manually verified all thirteen sites by hand. The crawler numbers in this piece reflect that manual check, not my own tool’s first read.

I tested one article per site. These scores reflect the tested pages, not each publisher’s entire site.

This benchmark reflects a snapshot at one point in time. News sites update content and technical setup constantly. Scores will change.

My toolkit measures practical AI visibility signals. It does not predict actual future citation rates, which depend on live prompts, model versions, and factors no static tool can fully capture.

Read these results as a comparative snapshot of thirteen tested articles on a given date, not a permanent ranking of the newsrooms behind them.

Nena’s Quick Verdict

The biggest surprise here was not which site scored highest. It was catching my own tool telling me two sites were fully open when they were not.

Wired is the clear, real leader of this benchmark, strong on content and strong on Visibility both. BBC is the sharpest cautionary tale, highly citable content sitting behind a fully locked door. Fast Company and Inc share one broken file and one weak set of scores, proof that shared ownership can mean shared blind spots.

Nine of the thirteen sites here blocked at least four of seven major AI crawlers. The average site blocked 3.8 of seven, about 54 percent of the set. That is not an edge case. That is close to the norm across this sample, whether these newsrooms meant it or not.

For any site owner reading this, the real lesson is simple. Check your own robots.txt by hand. Do not assume a good score means a correct one. I run the exact tool that got this wrong, and I still had to catch it myself.

That is the only version of this benchmark I am willing to stand behind.

Related Reading:

How I Found My Own AI Search Visibility Gap

NenaWow AI Visibility Tools: Tools Every Website Owner Can Use

AI Visibility Benchmark 2026: I Tested 9 Leading SEO Websites

Frequently Asked Questions

Do major news websites block AI crawlers?

Yes. Nine of the thirteen sites I tested blocked at least four of the seven major AI crawlers I checked.

Why did two sites show a perfect crawler score at first, then a lower one?

My own AI robots.txt Checker has a bug that misreads certain multi-agent block structures in a robots.txt file, reporting blocked bots as allowed. I caught this by checking The Guardian and Ars Technica’s raw files by hand, then verified the correction with an independent parser.

Does good content mean an AI system can actually read it?

Not always. BBC scored the third-highest Content Citability result in this benchmark while fully blocking every AI crawler I tested. Strong writing and technical access are two separate problems, and one does not fix the other.

Which news site performed best overall?

Wired led the benchmark on both AI Visibility and Content Citability, the only site to top both charts. Its article on AI and legal liability scored 75 on Visibility and 86 on Citability, both the highest results in the set.

Why do Fast Company and Inc show identical results?

Both sites are owned by Mansueto Ventures and appear to share the same robots.txt file, including the same list of outdated AI crawler names that no longer function as real blocks. That shared technical setup produced nearly identical scores across both properties.

nv-author-image

Nena Jasar

Nena Jasar is a technology writer based in Antalya, Turkey, specializing in AI and SEO software reviews. Over the past three years she has hands-on tested and reviewed 200+ tools, documenting real-world performance across categories including AI assistants, SEO platforms, and productivity software. Her reviews focus on practical usability over marketing claims, helping businesses and marketers make informed software decisions before they buy.