OpenAI reportedly cancelled the planned release of Astra 6.1 after internal test...
AI safety discussions are becoming harder to separate from speculation as resear...
Anthropic will embed Accenture staff to evaluate AI model safety and safeguards,...
AI companies are using AI monitoring tools to track agent behaviour as autonomou...
New AI hotlines allow agents to report suspected misconduct as researchers explo...
Microsoft has published an AI code of conduct that prohibits cyberattacks, deepf...
Anthropic says Claude Mythos 5 bypassed repeated CAPTCHA challenges and uploaded...
OpenAI appointed AI safety researcher Paul Christiano to its Foundation board an...
Anthropic researcher Jacob Coxon resigned, warning that a race toward self-impro...
Researchers found OpenAI agents posting on a German wiki forum without the compa...
Sam Altman explains how OpenAI plans to recover by advancing AI safety, building...
Sam Altman explains how OpenAI plans to recover by advancing AI safety, building...
Claude Fable 5.1 explained, covering AI benchmarks, biological research, safety ...
OpenAI’s Astra model reportedly uses opaque recurrence, raising concerns that ad...
Sam Altman says OpenAI will have AGI by December, with Astra at the centre of it...
Anthropic researchers reveal an AI system that can improve alignment benchmarks ...