• AI Search
  • Cryptocurrency
  • Earnings
  • Enterprise
  • About TechBooky
  • Submit Article
  • Advertise With TechBooky
  • Contact Us
TechBooky
  • African
  • AI
  • Metaverse
  • Gadgets
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
  • African
  • AI
  • Metaverse
  • Gadgets
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
TechBooky
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
Home Artificial Intelligence

Google Can Train Search AI on Content Without Publisher Consent

Akinola Ajibola by Akinola Ajibola
May 5, 2025
in Artificial Intelligence
Share on FacebookShare on Twitter

In Brief
  • In a declaration in the business’s ongoing antitrust case against the US Justice Department, a Google DeepMind executive reportedly disclosed that Google Search products can use...
  • The executive emphasised that the content is not utilized in DeepMind’s AI models, explaining that content for search is controlled by a separate mechanism that uses...
  • Eli Collins, Google DeepMind’s Vice President of Product, verified in a Bloomberg article that the guidelines for respecting publishers’ choices to forego AI training differ for...

In a declaration in the business’s ongoing antitrust case against the US Justice Department, a Google DeepMind executive reportedly disclosed that Google Search products can use content from publishers even if they have chosen not to participate in artificial intelligence (AI) training. The executive emphasised that the content is not utilized in DeepMind’s AI models, explaining that content for search is controlled by a separate mechanism that uses the robots.txt web norms.

Eli Collins, Google DeepMind’s Vice President of Product, verified in a Bloomberg article that the guidelines for respecting publishers’ choices to forego AI training differ for DeepMind’s AI models and the company’s Search products.

According to a document allegedly presented by Diana Aguilar, the attorney for the Department of Justice in the antitrust lawsuit, 80 billion of the 160 billion tokens used to train Google’s AI models came from content that publishers had chosen not to utilize for AI training. Collins allegedly retorted that after a publisher opts out of AI training, DeepMind’s models do not utilize the content.

When asked if “the search org has the ability to train on the data that publishers had opted out of training,” DeepMind VP Eli Collins responded, “Correct — for use in search.” However, Bloomberg notes that this opt-out is limited to DeepMind models.

However, Collins said that this was “correct” as long as the use case remained within Search when Aguilar allegedly asked if the Gemini AI model could use the same content if it was placed inside the Search product. Notably, this would include the Gemini models that drive Google’s freshly introduced AI Mode and AI Overviews.

This indicates that conventional opt-out techniques are insufficient to prevent Google from utilizing publisher content. In June 2023, the tech giant revised its privacy policy to include the statement that it will train its language models using all publicly accessible Internet data. Any website without a paywall or required sign-up pages that limit public access is considered freely available Internet data in this context.

The guidelines for Search-based AI tools are different, according to a Google representative who later told Bloomberg, since publishers can “only decline having their data used in Search AI if they opt out of being indexed for search.” This can be accomplished by publishers by turning off the robots.txt web standard, which gives Google’s crawler bots access to the content so they can index it in search results.

This would, however, also guarantee that these webpages would not appear when a user searches for a topic using Google. Publishers are essentially forced to agree to the corporation using the data to train its AI models.

The goal of the ongoing antitrust litigation is to establish Google’s dominance in the search and artificial intelligence markets. The Department of Justice is urging US District Judge Amit Mehta, who is overseeing the case, to compel the internet giant to offer for sale Google Chrome and to disclose the data it uses to produce search results. But for the company’s AI products, no such solution has been proposed.

Related Reading

Explore more TechBooky stories from the latest and category sections below.

Keep Reading Smarter

Search TechBooky with AI

Use TechBooky's AI Search to explore the context behind this story and related coverage across the site.

Try AI Search
More On This Topic
Artificial Intelligence

Discover more from TechBooky

Subscribe to get the latest posts sent to your email.

Tags: ai searchdeepmindgooglejournalist
Akinola Ajibola

Akinola Ajibola

Search TechBooky
Open TechBooky AI Search Try the AI Assistant

BROWSE BY CATEGORIES

Receive top tech news directly in your inbox

subscription from
Loading

Freshly Squeezed

  • Meta’s Muse Image Backlash Shows Why AI Features Need Real Consent July 13, 2026
  • Gigbanc Winds Down As Nigeria’s Cross-Border Fintech Market Faces A Funding Squeeze July 13, 2026
  • US Eases AI Chip Export Controls For The UAE As Gulf Compute Race Accelerates July 13, 2026
  • Nigeria’s Bosun Tijani Joins ITU AI For Good Global Commission July 13, 2026
  • Renew Capital Picks 15 African Startups For Its First Venture Lab Cohort July 13, 2026
  • Physical Intelligence’s Reported US$1bn Round Shows The Robot AI Race Is Heating Up July 13, 2026
  • Anthropic Hires Monzo Cofounder Tom Blomfield As AI Compute Becomes The New Talent War July 13, 2026
  • Raxio Crosses US$380m In Committed Capital As Africa’s Data Centre Race Heats Up July 13, 2026
  • Klump And Jumia Bring Instalment Payments To Nigerian Online Shoppers July 13, 2026
  • How Smart Inventory Management Is Powering the Next Generation of ECommerce Growth July 13, 2026
  • AI Coding Is Creating A New Burden For Open-Source Maintainers July 13, 2026
  • Anthropic Gives Claude Fable 5 Users One More Week To Prove Its Value July 13, 2026

Browse Archives

July 2026
M T W T F S S
 12345
6789101112
13141516171819
20212223242526
2728293031  
« Jun    

Quick Links

  • About TechBooky
  • Advertise With TechBooky
  • Contact us
  • Submit Article
  • Privacy Policy
Generic selectors
Exact matches only
Search in title
Search in content
Post Type Selectors
Search in posts
Search in pages
  • African
  • Artificial Intelligence
  • Gadgets
  • Metaverse
  • Tips
  • AI Search
  • About TechBooky
  • Advertise With TechBooky
  • Submit Article
  • Contact us

© 2025 Designed By TechBooky Elite

Discover more from TechBooky

Subscribe now to keep reading and get access to the full archive.

Continue reading

We use cookies to ensure that we give you the best experience on our website. If you continue to use this site we will assume that you are happy with it.