Safety monitoring and privacy usually pull in opposite directions. To catch misuse, a provider tends to keep records of what people do with its model, and those records are exactly what a bank or a hospital does not want sitting on someone else's servers. Anthropic says it has found a way to hold both at once, and the design is a quiet concession that its customers had a point.

The feature, announced on 1 September and called Enterprise Frontier Safeguards, lets business customers store Claude's activity and monitoring data in cloud infrastructure they control, rather than Anthropic's. That means Amazon S3, Azure Blob Storage or Google Cloud Storage, under the customer's own encryption keys and access policies. Automated systems still scan for serious misuse, but according to CNBC, no human review by Anthropic staff is required.

Why the change

This did not come out of nowhere. When Anthropic shipped Fable 5 in June, it introduced a 30-day data retention rule, and that policy landed badly with regulated industries that are not allowed to let sensitive records leave their control. As Unite.AI reported, the new safeguards are explicitly meant to replace that arrangement, pairing zero data retention on Anthropic's side with misuse detection that still works across sessions.

The company developed it in collaboration with more than a hundred customers spanning financial services, healthcare, manufacturing, telecoms, law, retail and the public sector. It also worked with the Analysis and Resilience Center for Systemic Risk, a group whose membership includes the chief information security officers of Goldman Sachs, Morgan Stanley, Citi, Bank of America and Wells Fargo. That is a telling list. These are the buyers who will not touch a model unless they can answer to their own regulators about where the data lives.

The tradeoff, stated plainly

There is a genuine tension worth naming rather than smoothing over. Detecting misuse of a powerful model is easier when the provider can see everything, and letting customers hold their own logs gives up some of that visibility in exchange for privacy. Anthropic's bet is that automated scanning against customer-held data preserves enough of the safety benefit while removing the reason enterprises balked. Whether that balance holds under real pressure is the thing to watch.

Anthropic said it will not charge for the safeguards, and that they will roll out in phases through the autumn. The move fits a pattern for the company, which has spent the year trying to make safety legible to the people who buy its products, from letting Claude take on some of its own alignment research to funding outside study of how AI affects the people who use it. Handing customers the keys to their own monitoring data is a smaller gesture, and a shrewd one. It removes an objection without softening the pitch.

Sources

  1. i. www.cnbc.com
  2. ii. www.unite.ai
  3. iii. www.marktechpost.com

Commentarii · 0

Add · a · Comment