AI Research35 min reading time

Claude Opus 5: Model Welfare

Don't Worry About the Vase
Read full post
Anthropic's Claude Opus 5 model achieved the highest scores on model welfare and alignment tests among recent models, though its performance may reflect strong test-taking abilities. The analysis highlights ongoing challenges and tradeoffs in improving AI model welfare as capabilities advance.

More in AI Research

Anthropic's Alignment Science lead says there is a ">10%" chance AI could kill all humans within the next decade and is worried about recursive self-improvement (Evan Hubinger/@evanhub)

Covered by 9 sources
AI Research4 min read

Suno trained its v6 AI music models with help from Warner and BMG

Covered by 5 sources

Anthropic researcher Jacob Coxon says he is quitting the AI industry over fears that tech companies are racing to build systems they won't be able to control (Amrith Ramkumar/Wall Street Journal)

Covered by 11 sources