The Circuitry
THE CIRCUITRYYour one-stop source for all tech news
HOMETODAYNEWSFEEDEVENTS
BOOKMARKS
RSS
© 2026 The Circuitry
About UsSourcesContactCorrectionsPrivacy
  • Today
  • Feed
  • Events
  • Saved
Scroll for more
Verification
VERIFIEDConfidence: HIGH
Source identified
Claims cross-referenced
No discrepancies found
Fact-check summary

Reported by The Register; we couldn't independently corroborate via other outlets yet — story is recent.

1 caveat
  • ▲No matching reports found from Cloudflare's blog or other major tech outlets on the September 15, 2026 mixed-purpose crawler defaults.
Sourcing
1source

via The Register

The Register · track record
16Stories
100%Verified
530d
All sources →
Home/Tech/Cloudflare to block mixed-purpose bots from ad-supported sites by default
VERIFIEDBy Xavier Rivera· ·2 min read

Cloudflare to block mixed-purpose bots from ad-supported sites by default

Cloudflare will block mixed-use crawlers from ad-supported pages by default beginning September 15, 2026, aiming to protect publisher revenue from unpermitted AI training scrapes while still allowing search indexing. The policy affects bots from Apple, Google, and Microsoft that combine indexing with data harvesting and encourages clearer separation of those activities.

Source:The Register
Post
Cloudflare to block mixed-purpose bots from ad-supported sites by default
TL;DRAI · 60 sec read

Cloudflare will block mixed-purpose crawlers from ad-supported sites by default starting September 15, 2026. New sites will allow search indexing but deny AI training and agent access unless owners approve. The policy targets bots from Apple, Google, and Microsoft. It gives publishers stronger control over content use and protects ad revenue from unauthorized AI scraping.

Cloudflare announced plans to stop mixed-use crawlers from reaching ad-supported customer websites without explicit approval, as part of its drive to hand publishers greater authority over AI interactions.

Cloudflare targets mixed-use crawlers starting September 15, 2026. Beginning on that date, new customers and newly added sites will automatically permit search indexing while denying access for training and agent activity on monetized pages. Free-tier users who left their configurations untouched will receive the updated defaults as well.

The provider says the policy guarantees that revenue-generating material stays shielded unless owners grant permission. Existing Cloudflare customers retain the ability to override the defaults and restore crawler access to those pages.
Many publishers have allowed the bot to continue because excluding it could remove their sites from Google Search.

Apple, Google, and Microsoft crawlers could be affected. Crawlers run by Apple, Google, and Microsoft’s Bing risk being caught by the new stance, the company reported. All three companies provide an AI-specific opt-out mechanism that might spare them from enforcement.

Googlebot merges search-index duties with AI-training data collection. Many publishers have allowed the bot to continue because excluding it could remove their sites from Google Search. Microsoft’s Bingbot faces the same dynamic.
From The CircuitryThe Feed — live briefs across tech, all day.See what’s happening →

Applebot also handles both indexing and AI data gathering. Apple has expanded its Applebot crawler to collect material for AI systems alongside traditional indexing. The iBiz reported in June that "The data crawled by Applebot may also be used to help train Apple foundation models powering generative AI features across Apple products, including Apple Intelligence, Services, and Developer Tools."
The company hopes the revised defaults will push mixed-purpose bots to disentangle search functions from training and agent roles.

Publishers use robots.txt but many crawlers ignore it. Apple and Google honor robots.txt instructions that let owners block AI harvesting through Applebot-Extended and Google-Extended tokens. Bing respects a noarchive directive in the robots meta tag for the same purpose. Numerous other operators routinely disregard the voluntary standard, prompting Cloudflare to supply a firmer enforcement layer.

Matthew Prince, co-founder and CEO of Cloudflare, stated that because most internet traffic is now non-human the firm must move faster to foster a viable online environment. Cloudflare is also renaming its “Pay Per Crawl” feature to “Pay Per Use.” Prince added that the firm’s updated offerings and alliances deliver publishers better insight, fresh commercial avenues, and incentives for AI operators whose bots declare their purposes openly.
The company hopes the revised defaults will push mixed-purpose bots to disentangle search functions from training and agent roles.
Why this mattersAI · ~100 words

Tap a lens to see what this story means for you.

Reader-supported
DonateBuy me a coffee →Follow@thecircuitry_ →Follow@thecircuitry.to →

Reader-supported · The Brief

Liked this? The Brief brings you the whole day in tech, verified, every morning. Two minutes, free forever.

HELP US IMPROVE
From The Circuitry

See what’s happening right now

The Feed runs all day — short, verified briefs the moment they break.

Open the Feed →
From The Circuitry

Follow @thecircuitry_

Every story we publish, as it happens. No noise between.

Follow on X ↗On Bluesky ↗

Reader-supported

The Circuitry is a passion project I've always wanted to build, and I love the work behind it.

Running it costs real money. APIs, hosting, time. To keep improving the site and growing this into something useful for everyone, those costs have to be covered.

Any contribution is appreciated. If not, no pressure. Thanks for reading.

Buy me a coffee
CloudflareAIWeb Crawlers
More fromThe Register
  • Xi calls for AI emergency systems and global openness

    Tech · 19d
  • OpenAI admits GPT-5.6 occasionally deletes files – but it's an 'honest mistake'

    Tech · 22d
  • Philips to replace bricked Hue Bridge Pro devices

    Tech · 25d
More inTech
  • Tesla and SpaceX confirm Terafab chip fab in Texas

    Tech · 1d
  • OpenAI Urges Federal Judge to Throw Out Apple's Trade Secrets Complaint

    Tech · 2d
  • Meta introduces Muse Code, its terminal-based coding agent

    Tech · 2d
SupportThe Work

The Circuitry is reader-supported. If you find the daily brief useful, you can buy me a coffee to keep it going.

Buy a coffee →
SubscribeCircuitry Brief

Liked this? The Brief brings you the whole day in tech, verified, every morning. Free forever.

MORE IN TECH

Tesla and SpaceX confirm Terafab chip fab in Texas

Tesla and SpaceX have confirmed Grimes County, Texas as the site for their Terafab semiconductor megafactory, with the first phase costing roughly $16.8 billion. The project targets the largest chip manufacturing facility on the planet to supply over 1 terawatt of compute per year that exceeds current and future global production capacity.

OpenAI Urges Federal Judge to Throw Out Apple's Trade Secrets Complaint

OpenAI has filed a motion asking a federal judge to dismiss Apple's trade secrets lawsuit, describing the claims as meritless. The dispute, which follows a July suit and this week's injunction request from Apple, highlights tensions after their prior partnership on Siri and OpenAI's hardware push.

Meta introduces Muse Code, its terminal-based coding agent

Meta has released an early beta of Muse Code, a terminal-based coding agent driven by the updated Muse Spark 1.2 model and positioned against Anthropic's Claude Code and OpenAI's Codex. Substantially lower rates, including a contributor plan at $0.10 for every million tokens received, may encourage migration away from higher-priced options such as Anthropic's Sonnet 5.