Cloudflare's September 15 default also blocks Googlebot
Ricardo Argüello, September 14, 2026
CEO & Founder
General summary
On September 15, 2026 Cloudflare changes its defaults for AI traffic, blocking Training and Agent crawlers on ad-supported pages. Multi-purpose crawlers resolve by the most restrictive matching rule, so Googlebot, Applebot and BingBot get caught by a Training block.
- Announced July 1, 2026; defaults take effect September 15 for new domains, new customers and every existing free-tier customer
- Training and Agent blocked by default on pages that display ads, Search still allowed
- Googlebot, Applebot and BingBot are multi-purpose, so a Training block removes them too
- Content Signals adds search, ai-input and ai-train directives to robots.txt, portable off Cloudflare
- Pay Per Crawl becomes Pay Per Use, paying publishers when content is used in an answer rather than when a page is fetched
Picture a gate with three keys: one for the postman, one for the courier bringing a package, one for the truck hauling boxes to a warehouse. Cloudflare just installed that gate on your site. The catch is that Google's postman also drives the truck, and taking away the truck key takes both.
AI-generated summary
There is a checkbox in the Cloudflare dashboard that says block AI training crawlers.
Tick it and you also remove Googlebot, Applebot and BingBot. Your site leaves the search index.
That is not a bug and it is not a trick. It is the only coherent answer Cloudflare can give when one crawler does two jobs and you allow one of them while forbidding the other. But it means a content decision and a search visibility decision now share a single control, and on September 15 that control starts moving on its own for a large slice of the web.
What actually flips on September 15
Cloudflare announced this on July 1, splitting automated traffic three ways.
Search indexes your content so it can answer questions about it later, and sends referral traffic back. Agent is automated activity acting in real time for a person, the fetch bot that pulls a page because someone asked a chatbot a question. Training takes your content to train or fine-tune a model.
From September 15, on pages that display ads, Training and Agent are blocked by default. Search stays open.
The new defaults reach new domains onboarding to Cloudflare, new customers, and every existing customer on the free tier. Paid customers with configured sites keep the rules they already set.
Matthew Prince framed it as an ecosystem problem: “Now that the majority of traffic on the Internet is non-human, we must go further and act faster so that a sustainable ecosystem can emerge.”
The collision nobody budgeted for
Cloudflare says multi-purpose crawlers, the ones combining Search with Training, are allowed or blocked according to all of their behaviors, and that defaults are enforced by the most restrictive applicable rule.
Then it names the three that matter. Googlebot. Applebot. BingBot.
So the strictest rule wins, and the strictest rule is the one you set to protect your corpus. An infrastructure engineer ticks a box on a Tuesday afternoon, marketing finds out five weeks later when impressions have already fallen off, and by then nobody remembers the change was made.
Cloudflare’s own community forum has threads from people hitting this before the deadline. The mechanism is documented. The surprise is entirely on the operator side.
Content Signals, and a different billing event
Alongside the categories, Cloudflare extended Content Signals, which already existed: three machine-readable directives in robots.txt. search for building an index, ai-input for feeding real-time answers, ai-train for training or fine-tuning. Same taxonomy, expressed in a file you own whether or not you use Cloudflare. That portability is worth more than the dashboard toggle.
Pay Per Crawl also became Pay Per Use. The old model charged bots per fetch. The new one pays publishers when their content is used inside a generated answer. Ceramic.ai and You.com are the named launch partners.
Moving the billing event from the request to the use is the right direction, and it changes what you have to measure. Fetches are countable. Citations are not, at least not the way most teams try. We wrote about that when Muck Rack found that the journalists PR teams pitch and the ones AI ends up citing overlap by just 2%, and about why identical prompts return different sources in AI visibility measurement and sampling noise.
Set the policy, do not inherit it
Blocking Training while allowing Agent is a defensible position. You keep your corpus out of the next training run and you still show up when a person asks a chatbot about you. Blocking everything is also defensible, and it costs visibility. Both are real choices.
Inheriting a default is not a choice.
So before Tuesday, check which of your domains sit on the free tier, confirm whether your pages count as ad-supported under Cloudflare’s definition (a remarketing pixel is easy to forget), and if anyone has already enabled a Training block, open Search Console today and verify Googlebot is still crawling. Do not wait for the traffic graph to tell you.
At IQ Source this is the first hour of an AEO engagement: decide the policy per category, write it into robots.txt and Cloudflare so both agree, then stand up measurement that survives the sampling noise. The date is fixed either way. What is still open is who in your company gets to make the call.
Set your crawler policy before TuesdayFrequently Asked Questions
Cloudflare's defaults change. On pages that display advertising, crawlers in the Training and Agent categories are blocked without any configuration, while Search stays allowed. The new defaults apply to new domains onboarding, new customers, and all existing customers on Cloudflare's free tier.
Yes. Cloudflare treats Googlebot, Applebot and BingBot as multi-purpose crawlers because they combine search indexing with training, and it resolves them by the most restrictive matching rule. Turning on a Training block removes all three, which takes the site out of the search index.
Search covers crawlers that index content to answer questions later, with referral traffic in return. Agent covers automated activity acting in real time on a person's behalf, such as chat fetch bots. Training covers crawlers taking content to train or fine-tune a model.
Pay Per Crawl charged AI bots for fetching a page. Pay Per Use moves the billing event: the publisher is paid when the content is actually used inside a generated answer, not when the request happens. Cloudflare named Ceramic.ai and You.com as launch partners.
Related Articles
85.5% Earned Media, 2% Overlap With What You Pitch
The stat every PR team is sharing and the 2% overlap between pitches and AI citations come from different measurements. Here is what each one is good for.
That AI Visibility Percentage Your Vendor Sold You May Be Noise
Profound and Evertune are publicly arguing over how to measure whether your brand shows up in ChatGPT. Most visibility numbers never say how they sampled.
Cloudflare's x402 Toll Booth: 15 of 15 Facilitators Failed
Cloudflare, AWS, Visa, and 37 other companies just built payment rails for AI agents. A USENIX paper found security violations in all 15 facilitators tested.
Pew Measured It: A Third of the New Web Is AI-Written
Pew ran 490,000 pages through a detector. The useful part is the list of stylistic markers it published, because you can audit your own archive against it.
Your Shorts are deposits into the AI that cites you
Gary Vaynerchuk says YouTube Shorts became his number one platform. Not for the views, but because every video is a deposit into the AEO fight ahead.
Your marketing asks AI what to cut. Ask this instead.
Box created 13 new AI roles, one to market to industries it could not staff before. The question is not what AI lets you cut, but what it makes possible.