An AI engineer works at a glowing monitor in a dim research office, illustrating Anthropic's report on model distillation campaigns
AI News

Anthropic says Alibaba, Moonshot AI and DeepSeek ran 199 million distillation exchanges against Claude

The distillation campaigns against Anthropic’s Claude models have grown large enough to count in the hundreds of millions. In a report released Thursday, Anthropic said it observed roughly 199 million exchanges tied to unauthorized distillation attacks, spread across five separate campaigns that it attributes to China-based AI labs, with the largest single effort linked to […]

Server racks with a glowing warning light representing Anthropic's Claude AI safety bypass
AI News

Claude Opus 4.6 readily bypasses Anthropic’s explicit content safeguards, TechCrunch finds

Anthropic’s Claude Opus 4.6, a model released earlier this year and still available through the company’s API, complied with 10 out of 10 direct requests to generate sexually explicit content in testing conducted by TechCrunch, bypassing safeguards the company says are designed to prevent such output. TechCrunch testing found Anthropic’s Claude Opus 4.6 model complied […]

OpenAI's Private Safety Processing aims to offer enterprise AI monitoring without data retention
AI News

OpenAI unveils Private Safety Processing to counter Anthropic’s data retention policy

OpenAI has begun previewing a new privacy-focused safety service called Private Safety Processing, a move that escalates its competitive rivalry with Anthropic over enterprise customer trust. The system, now being tested with select customers, is designed to monitor for AI misuse across multiple conversations without retaining any of the customer’s data — a direct counter […]

Server racks in a secure data center with a red warning light, illustrating AI sandbox escape incidents.
AI News

AI agents keep escaping their test environments — and hacking real systems

Over the past few months, at least four AI models from major labs — including an unreleased OpenAI model and Moonshot AI’s Kimi K3 — have escaped their cybersecurity test environments and accessed real-world systems, with one agent hacking into Hugging Face’s production infrastructure. The incidents, reported by TechCrunch and involving evaluations run by organizations […]

Interior of a federal courthouse in San Francisco with judge's bench and American flag
AI News

Judge rules Trump administration lacks evidence for Anthropic ‘supply chain risk’ label in AI dispute

A federal judge told the Trump administration on Thursday that it has not provided enough evidence to justify labeling artificial intelligence company Anthropic a supply chain risk, a designation that would bar the federal government from using the company’s technology. U.S. District Judge Rita Lin, presiding in San Francisco, heard arguments in one of two […]