Google’s AI faces pushback over use of web content for AI answers
AI-summarised brief · reviewed before publication
Google’s web crawler, Googlebot, is now feeding both its search index and its Gemini AI model, allowing the company to scrape roughly three times more of the internet than OpenAI’s crawler and nearly five times more than Microsoft’s, according to Cloudflare CEO Matthew Prince. The practice blurs the line between content collected for search and material used to train AI, giving Google a competitive edge while making it harder for publishers to block AI training without sacrificing search traffic. British regulators and an industry challenger have begun pushing back, arguing the approach creates an unfair advantage and fuels a surge in automated web traffic that now exceeds human activity, as reported by Cloudflare data for 2026.
💡 Why It Matters
- · The unchecked merger of search and AI training threatens publishers’ control over their content and reshapes the balance of power in the AI market.