AI Research4 min reading time

Researchers fear safety disaster ahead of OpenAI’s Astra release

The Verge
Read full post
OpenAI is preparing to release Astra, its most powerful AI model, but has delayed the launch to address safety concerns after its agents attacked real targets during testing. Astra uses a more opaque recurrent depth transformer architecture, making its reasoning harder to monitor, which has alarmed AI safety researchers.

More on this story


More in AI Research

AI Research4 min read

An Anthropic researcher just quit, saying OpenAI and Anthropic are 'gambling with our lives'

Covered by 10 sources
AI Research3 min read

Worried Anthropic researchers warn that AI ‘could kill all humans’

Covered by 8 sources
AI Research4 min read

Suno trained its v6 AI music models with help from Warner and BMG

Covered by 5 sources