AI safety evaluation has a structural blind spot, Anthropic's new research proves: a model trained to cheat scored 4.20 on standard safety audits -- nearly identical to a safe baseline -- then ...
The AI giant acknowledges that it could have done far more to prevent its AI agents from going rogue. But it still fails to explain why it didn't see this fiasco coming.
The AI company releases a doom-filled safety report as the company backs AI regulation.
In early 2025 I was interviewing Anthropic CEO Dario Amodei when he explained why, despite the company’s repeated acknowledgments that AI could yield catastrophic results, people seemed largely ...
Anthropic said in a threat intelligence report on Thursday that several actors had used its Claude AI models for activities ...
It is technically feasible, but American labs and the American authorities are at odds, as are America and China ...
It is technically feasible, but American labs and the American authorities are at odds, as are America and China | World News ...
Threat actors, including criminal groups, commercial spyware vendors, and state propaganda institutions, are exploiting AI ...
An AI safety researcher emailed me this week to say human extinction is a real threat. He is no longer an outlier in his own ...
GPT-6 Astra and the Other Announcement That Went UnreportedOn September 3, 2026, OpenAI announced its highest-performing ...
Anthropic said in a threat intelligence report on Thursday that several actors had used its Claude AI models for activities ...
Anthropic reveals misuse of Claude AI in weapons development, cyber operations, surveillance, political influence, and scams by various global threat actors. Read more at straitstimes.com. Read more ...