cloudflare

Content Independence Day, one year on- building the business model for the agentic Internet (opens in new tab)

Cloudflare argues that generative AI has rapidly replaced the traditional web model in which publishers traded content access for search referrals. With AI now driving much of online discovery and crawler activity, content is increasingly consumed without users visiting its source. The company says a new market is emerging in which transparency, access controls, scarcity, and licensing can help publishers regain economic value.

AI’s rapid transformation of the Internet

  • Generative AI adoption has reached more than 2.5 billion regular users—over 30% of humanity—in roughly 3.5 years, reportedly more than twice the adoption speed of smartphones.
  • Users now spend only about 15 minutes on the open web for every hour spent searching for information.
  • Instead of visiting and comparing multiple websites, users increasingly receive consolidated answers directly from AI systems.
  • More than 50% of Internet traffic is now non-human, marking the arrival of what Cloudflare calls the “agentic Internet.”

Crawlers are increasingly focused on AI

  • AI training accounted for 52% of crawler requests in June 2026, up from 22% in spring 2025.
  • Mixed-use crawlers, combining search, agent activity, and training, represented more than 36% of crawler traffic.
  • Traditional search crawlers make up a smaller share of activity, even though they remain important for sending visitors to publishers.
  • Mixed-purpose crawling makes it difficult for site owners to remain visible to AI-driven discovery without also giving away content for training without compensation.

The traditional web business model is breaking down

  • Historically, publishers allowed search engines to crawl their content in exchange for visibility and referral traffic.
  • AI systems now answer questions, conduct research, compare products, and complete tasks without necessarily sending users to original sources.
  • Content can therefore be crawled, indexed, and monetized by AI companies while the original publisher receives little or no traffic.
  • News and media organizations experienced the disruption first, but retail, software, IT, finance, and other sectors are also affected.
  • Some heavily crawled categories have seen human traffic fall by as much as 40% in under a year.
  • Publishers are preparing for “Google Zero,” in which search referrals provide little meaningful traffic.

The impact extends across industries

  • Any organization publishing proprietary information online may need a strategy for AI access and monetization.
  • The issue affects not only traditional publishers but also businesses whose websites contain valuable product, technical, financial, or industry knowledge.
  • Cloudflare frames the sustainability of online content as an economic and public-interest concern because the Internet remains a major global information resource.

Building a market for content

Cloudflare says Content Independence Day focused on three goals:

  • Give site owners transparency and control over how their content is accessed and monetized.
  • Create scarcity by allowing publishers to restrict or selectively permit AI access.
  • Establish a marketplace where publishers and AI companies can discover, license, and price content.

According to the post, these efforts have helped create the early conditions for a monetized content market.

Control and data create negotiating power

  • Cloudflare’s attribution, business intelligence, and enforcement tools let publishers observe AI access at the network level.
  • These tools provide stronger practical enforcement than voluntary mechanisms such as robots.txt.
  • Publishers can identify:
    • How often LLMs attempt to access their content
    • Which competing AI systems are crawling their sites
    • Which URLs are most in demand
    • The relationship between crawling and referrals
  • Restricting or controlling access creates scarcity, which gives publishers leverage in licensing negotiations.
  • Better operational data reduces information asymmetry and allows content owners to negotiate with evidence rather than guesswork.

Ultimately, the post recommends treating online content as an economic asset rather than an unlimited free input. Publishers should measure AI consumption, control access, and pursue licensing arrangements so that the agentic Internet can support content creation instead of undermining it.