Articles  /  For Publishers

For Publishers

AI Crawlers and Publishers: Block Them, or Profit From Them?

AI bots download content from media at an enormous ratio to the visits they return – with some it is tens of thousands of downloaded pages per visit. A breakdown of the options a publisher realistically has today, and why blanket blocking need not be the best answer.

Redakce Linketica · 8. 8. 2026 · 9 min čtení
Poměr stažených stránek k vráceným návštěvám u AI botů
AI boti berou hodně a vracejí málo – poměr se u jednotlivých platforem liší řádově.

For publishers it is the most uncomfortable topic of the past two years. Language models learn and answer from content somebody wrote and paid for – and the traffic that used to come back for it is shrinking. Yet the question “to block or not” has no single correct answer; it has several options with different consequences.

How unequal the ratio is

Cloudflare’s 2025 data gives a frame worth remembering: while with Google roughly 14 downloaded pages fall on one referred visit, with Perplexity it is in the hundreds, with OpenAI approximately a thousand and with Anthropic tens of thousands. In other words: the classic search engine was a trade; this, so far, is mostly a one-way take.

Add the data on reader behaviour: according to Pew Research, with an AI overview shown, a user clicks a classic result in only 8% of cases versus 15% without it – and a link inside the overview in only 1%. Zero-click hits media harder than anyone else, because their product is exactly the information the overview provides.

Distinguish whom you are actually dealing with

A blanket block of “AI bots” is a blunt tool, because they are not one group:

Each of them can be controlled separately in robots.txt. A sensible default position for most media: consider the training bots, let the search and on-demand ones through. Removing yourself from the answers with citations means disappearing from the place where part of the audience today forms its idea of which media are worth reading.

Three strategies that make sense today

  1. Selective access. Open up to search bots, restrict the training ones – especially for the archive and exclusive content. Low risk, no loss of visibility.
  2. Paid access. Infrastructure players (Cloudflare and others) today offer modes where the crawler either pays or does not get through. For a medium with a solid volume of content it is a real path to turning the take into income.
  3. Licensing. Direct deals with model operators. Available mainly to large publishers, but smaller titles can act jointly within industry associations.

Why being a cited source pays

Traffic is not the only currency. When your medium repeatedly appears as a source in AI answers, its perceived authority grows – and that is exactly what advertisers buy from you. A title the models refer to is more valuable to PR article buyers than a title with the same traffic that is missing from the answers. Citability can even be documented: put queries from your field to the models and record how often you are listed as a source. It is an argument for a sales negotiation that few use so far.

What to do this week

  1. Look into your logs at how much traffic the AI bots really take from you – deciding without numbers goes badly.
  2. Go through robots.txt and split the bots into the three groups above instead of one blanket rule.
  3. Verify whether the models mention you on the topics you build your brand on.
  4. Add this figure to your advertiser deck next to traffic and DA.

And then it is only about the buyers who appreciate this learning about your title. That is exactly why it makes sense to have the medium visible where advertisers and agencies pick media – for example on the Linketica marketplace, where you manage the parameters, rules and price list yourself.

Médium jako citovaný zdroj v AI odpovědích
Být citovaným zdrojem má hodnotu i bez prokliku – pro čtenáře i pro inzerenty.

Časté otázky

How much content do AI bots download compared with the visits they return?
According to 2025 Cloudflare data, with Google about 14 downloaded pages fall on one referred visit, with Perplexity hundreds, with OpenAI roughly a thousand and with Anthropic tens of thousands.
Should a publisher block AI crawlers across the board?
A blanket block is a blunt tool. More sensible is to distinguish training bots (consider restricting), search bots with citations and bots fetching a page on demand – mostly let the latter two through, otherwise you disappear from the answers where the reputation of media is formed.
What good is AI citability to a medium when it brings no clicks?
It raises the title’s perceived authority, which is exactly what advertisers buy from a publisher. A title cited by the models is more valuable to PR article buyers than a title with the same traffic that is missing from the answers.

Zdroje a doporučená literatura

🤖

Zjistěte zdarma, jak vás vidí AI

Nechte jméno, e-mail a telefon a odemkneme vám online audit AI viditelnosti + checklist „Doporučuje vás AI?“ ke stažení.

Chcete rovnou nakupovat PR články a zmínky? Vyzkoušet Linketicu  ·  Nezávazná poptávka na míru

Zadáním kontaktu souhlasíte s jeho použitím pro zpřístupnění auditu a zaslání tipů. Údaje nepředáváme dál. · Linketica.com