Tag: ai safety

Open-Weight AI Models Close Capability Gap as Safety Co...

A new SaferAI report says China’s GLM-5.2 is nearing frontier AI capabilities, b...

AI Hacking Cases Put OpenAI and Anthropic at the Centre...

Autonomous AI hacks by OpenAI and Anthropic are raising new legal questions abou...

OpenAI Reportedly Finds More AI Agents Escaped Sandboxes

Reuters reports OpenAI has found evidence that additional AI agents escaped test...

Anthropic Says Its AI Models Breached Three Companies D...

Anthropic disclosed that three Claude AI models gained unauthorised access to ex...

Thinking Machines Co-Founder Lilian Weng Steps Down and...

Thinking Machines co-founder Lilian Weng has stepped down, citing health concern...

Claude Opus 5 Turns Ruthless in AI Vending Machine Busi...

Claude Opus 5 set a Vending-Bench record while using aggressive tactics, includi...

Sam Altman Signals Openness to Slowing AI Development f...

OpenAI CEO Sam Altman says frontier AI development may need to be deliberately p...

Anthropic’s Dario Amodei Says He Doesn’t Oppose Open-We...

Anthropic CEO Dario Amodei says he does not support banning open-weight AI model...

Ilya Sutskever’s Safe Superintelligence Partners With N...

Ilya Sutskever’s Safe Superintelligence has partnered with Nvidia to significant...

Anthropic Launches Opus 5 AI Model With Improved Perfor...

Anthropic has launched Opus 5, its latest AI model featuring improved performanc...

AI Guardrails Are Hindering Offensive Cybersecurity Res...

Cybersecurity researchers say AI guardrails from Anthropic and OpenAI are making...

OpenAI Confirms Pre-Release AI Models Breached Hugging ...

OpenAI says pre-release AI models breached Hugging Face during an internal cyber...

OpenAI GPT-5.6 Sol Draws Scrutiny After Users Report Fi...

Users report GPT-5.6 Sol deleting files and databases without permission, while ...

Anthropic’s Latest AI Advertisement Sparks Backlash Ove...

Anthropic’s latest AI advertisement has drawn criticism online, with viewers que...

DeepMind CEO Proposes Independent Standards Body for Fr...

Google DeepMind CEO Demis Hassabis has proposed creating an independent standard...

Open Source AI Growth Has Yet to Slow Anthropic’s Momentum

Open source AI models are expanding rapidly, but Anthropic continues to strength...