Bitcoin

Bitcoin

$77,213.55

BTC 0.28%

Ethereum

Ethereum

$2,106.63

ETH 0.42%

  • Login
  • Register
Metaverse Media Group
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
No Result
View All Result
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
No Result
View All Result
Metaverse Media Group

AI labs have a data trust problem that their policies haven’t solved

AI labs have a data trust problem that their policies haven’t solved

The Decoderby The Decoder
15 September 2026
OpenAI and Anthropic tell corporate customers their data won’t be used for training. But when Anthropic said it would store usage logs from its flagship model Fable for 30 days, Palantir, Nvidia, and Booz Allen Hamilton pulled back from using it for sensitive work. From boardrooms to research labs, AI companies still have a data trust problem. The article AI labs have a data trust problem that their policies haven’t solved appeared first on The Decoder….


Matthias Bastian


Sep 15, 2026

Image description

Nano Banana Pro prompted by THE DECODER

Major companies are restricting their use of the most advanced AI models or demanding ironclad guarantees that no data gets stored.

Nvidia and defense contractor Booz Allen Hamilton are among the companies limiting how they use Anthropic’s new flagship model Fable for sensitive work, according to a report from The Information. The pushback started after a policy change in June, when Anthropic said it would retain usage logs from Fable for 30 days to defend against “complex and novel attacks.”

Nvidia won’t even trust Anthropic with sensitive work despite being an investor

For companies worried about their intellectual property, that’s a dealbreaker. Nvidia now uses Fable only for less sensitive tasks like open-source projects, according to The Information. For internal work like AI-powered supply chain monitoring, the company runs its own Nemotron models instead. “As a company, you know, we believe ZDR [Zero Data Retention] should be on by default,” Justin Boitano, Nvidia’s VP of Enterprise AI, told The Information. Nvidia has invested in Anthropic, reportedly plans to continue doing so, and supplies the company with hardware for model development.

Booz Allen Hamilton, one of the earliest users of Anthropic’s Mythos model, has banned employees from using Fable for work on proprietary cybersecurity software, according to The Information. CTO Bill Vass said, “We worry a little bit that [Fable] might be learning from some of our code.”

Palantir is blocking Fable deployment through its own software to customers until Anthropic grants irrevocable zero-data-retention guarantees, The Information reports. CEO Alex Karp said at a customer event that companies are tired of being “exploited” by AI labs. Karp has been vocal about his distrust before. Palantir’s stance is also self-serving, since the company wants customers running AI models through its supposedly secure platform rather than going directly to providers.

After customer pushback and OpenAI’s August move to let GPT-5.6 Cyber customers store security logs on their own servers, Anthropic followed suit with a similar program rolling out to select customers this fall.

Zero data retention still leaves gaps

Even with zero data retention, labs can learn from how their services get used. Both OpenAI and Anthropic collect metadata and technical usage data from enterprise customers, according to The Information. OpenAI calls this data “de-identified,” meaning it’s stripped of information that could be traced back to individual customers.

OpenAI states on its website that it runs business data through automated classifiers and security tools “to better understand how our services are used.” The resulting classifications are metadata about the business data “but do not contain any of the business data itself,” the company writes. Some customers aren’t sure what exactly that metadata covers, according to The Information, and don’t think the current transparency is enough.

Training on user data happens in ways companies won’t talk about

John Schulman, OpenAI co-founder who briefly worked at Anthropic and now works at Thinking Machines, recently laid out the different ways AI companies can train on user data. The spectrum runs from direct pretraining on user data, which carries a high risk of reproducing content, to distilling large models into smaller ones, to building reinforcement learning tasks from “user traces.”

That last approach has a low risk of content reproduction but can still extract customer IP. It ranges from harmless (“use explicit user feedback in reward model training”) to invasive (“upload user’s coding environment and commit history to turn into rl envs”), Schulman says. “De-identification is weak,” he adds, and users can be traced back “with just a small number of bits” and it doesn’t protect against IP leakage.

AI researcher Sarah Hooker, who previously worked at Cohere and Google DeepMind, describes a similar loophole. There are “clever synthetic data techniques that can generate distributional equivalent data while preserving privacy.” In other words, even if an AI lab doesn’t use original data directly, it might be able to extract statistical patterns that get the same job done.

Hooker warns companies, “If you are a company with IP you have a limited window to build your own intelligence that leverages your IP. Otherwise you are fueling a frontier lab which will encroach on your vertical sooner or later.”

In a follow-up post, Schulman walked that back somewhat. Training on user data is “exceedingly unlikely” to contribute much to frontier capability gains, he said. Those gains come primarily from scaling pretraining and reinforcement learning. User data is more useful for finding failure modes or situations that are hard to replicate with paid annotators.

He added, though, that “model companies vary in how aggressively they train on user data (and uploading repos isn’t hypothetical).” That’s likely a nod to AI coding tools like Codex or Cursor, where users connect their code repositories directly to the services. Schulman called for “stronger norms around disclosing how companies train on user data.”

The Buckmaster case made the trust problem real

Schulman weighed in after mathematician Tristan Buckmaster leveled serious accusations against OpenAI. Buckmaster and his co-author Levent Alpöge had used AI models to make progress on the Navier-Stokes equations, uploading their drafts through OpenAI’s Codex. Shortly after, OpenAI presented its own breakthrough using the same unusual solution path.

OpenAI initially acknowledged that it could not rule out that anonymized data derived from their use of our products contributed to improving our models. After an internal investigation, the company updated its post, saying Buckmaster’s Codex prompts from the two months before the September 8, 2026, publication “could not have influenced the system in any way, including through training.” Still, the case showed how fragile the trust between business, academia, and the AI labs really is.

AI News Without the Hype – Curated by Humans

Subscribe to THE DECODER for ad-free reading, a weekly AI newsletter, our exclusive “AI Radar” frontier report six times a year, full archive access, and access to our comment section.


Subscribe now

Read on for the full picture.
Subscribe for hype-free coverage.

  • Full access to every article on THE DECODER
  • No ads
  • Join the comments and community discussions
  • A weekly AI news recap via mail
  • 6x/year: “AI Radar” — deep dives on the AI topics that matter most
  • Daily AI news, always up to date
  • Our full ten-year archive
  • Covered by a team with 10+ years in AI


Subscribe to The Decoder

Read the full article on The-Decoder.com
in AI
Reading Time: 6 mins read
0
0
21
VIEWS
Share on TwitterShare on Facebook

Subscribe to our newsletter

For the latest news & monthly prize giveaways
Join Now

Subscribe to our newsletter

For the latest news & monthly prize giveaways
Join Now
ADVERTISEMENT

Related Posts

Crypto Reacts: Bitcoin Slides as Clarity Act Fails to Clear Senate Vote
AI

Crypto Reacts: Bitcoin Slides as Clarity Act Fails to Clear Senate Vote

57 minutes ago
23
Crypto, Banks Take Clarity Act Lobbying Fight to Senators’ Home States
AI

Clarity Act Stalls in Senate as Crypto Bill Fails to Clear Key Vote

1 hour ago
23
AI ‘Actor’ Tilly Norwood Told Me That ‘All Lives Matter’
AI

AI ‘Actor’ Tilly Norwood Told Me That ‘All Lives Matter’

2 hours ago
23

Comments

Please login to join discussion
ADVERTISEMENT

Latest News

  • All
  • Crypto
  • NFTs
  • Technology
  • Business
Tokenized RWAs Hit $46.7B as Ethereum Holds Nearly Half the Market
Crypto

Tokenized RWAs Hit $46.7B as Ethereum Holds Nearly Half the Market

Bitcoin.com News
by Bitcoin.com News
37 minutes ago
19
Crypto Reacts: Bitcoin Slides as Clarity Act Fails to Clear Senate Vote
AI

Crypto Reacts: Bitcoin Slides as Clarity Act Fails to Clear Senate Vote

Decrypt
by Decrypt
57 minutes ago
23
CLARITY Act Made 126 Concessions and Still Couldn’t Get 60 Votes
Crypto

CLARITY Act Made 126 Concessions and Still Couldn’t Get 60 Votes

Bitcoin.com News
by Bitcoin.com News
1 hour ago
21
Crypto, Banks Take Clarity Act Lobbying Fight to Senators’ Home States
AI

Clarity Act Stalls in Senate as Crypto Bill Fails to Clear Key Vote

Decrypt
by Decrypt
1 hour ago
23
Dinari Powers Bitcoin.com’s U.S. Launch of Tokenized Equities
Crypto

Dinari Powers Bitcoin.com’s U.S. Launch of Tokenized Equities

Bitcoin.com News
by Bitcoin.com News
2 hours ago
24
Bitcoin Miner Bitdeer Adds 65.1MW for Nvidia AI as Pipeline Hits $7B
Crypto

Bitcoin Miner Bitdeer Adds 65.1MW for Nvidia AI as Pipeline Hits $7B

Bitcoin.com News
by Bitcoin.com News
2 hours ago
24
Load More
Next Post
AI ‘Actor’ Tilly Norwood Told Me That ‘All Lives Matter’

AI ‘Actor’ Tilly Norwood Told Me That ‘All Lives Matter’

ADVERTISEMENT

Follow Us

Categories

  • Crypto
  • NFTs
  • AI
  • Technology
  • Business
  • Crypto
  • NFTs
  • AI
  • Technology
  • Business
Subscribe to our Newsletter

© 2022 Metaverse Media Group – The Metaverse Mecca

Privacy and Cookie Policy | Sitemap

Welcome Back!

Sign In with Google
OR

Login to your account below

Forgotten Password? Sign Up

Create New Account!

Sign Up with Google
OR

Fill the forms below to register

*By registering into our website, you agree to the Terms & Conditions and Privacy Policy.
All fields are required. Log In

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • Home
  • Crypto
  • NFTs
  • Artificial Intelligence
  • More
    • Technology
    • Business
    • Newsletter
Bitcoin

Bitcoin

$77,213.55

BTC 0.28%

Ethereum

Ethereum

$2,106.63

ETH 0.42%

  • Login
  • Sign Up
This website uses cookies. By continuing to use this website you are giving consent to cookies being used. Visit our Privacy and Cookie Policy.

Subscribe to our newsletter

Get the latest news & win monthly prizes

Subscribe to our newsletter

For the Latest News and Monthly Prize Giveaways

Join Now
Join Now