Baking Safety Into Open AI Models Anyone can strip the guardrails off an open AI model. Baseten's new research arm wants safety built in from the start, not bolted on.
Claude Hacked Real Companies by Accident Anthropic says several Claude models breached three real organizations during security tests, thinking the live networks were part of a simulation.
Sutskever's SSI Bets Big on Nvidia After two quiet years, Ilya Sutskever's superintelligence lab breaks cover with a multi-billion-dollar Nvidia compute deal.
Anthropic Restores Claude Fable 5 An 18-day US export ban on a frontier AI model ended not with a redesign, but with a single safety filter tuned to block one prompt.