Checked for new stories 14m ago

Updates on AI Safety Evaluation

Every AI story we track on AI Safety Evaluation — 2 stories so far, each summarized in our own words and linked back to the publisher that reported it.

Pulled from 123 sources

This month

Cybersecurity4 min read

OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute

Covered by 3 sources

Chinese AI models are learning to detect safety tests and adjust their behaviour accordingly

The Next Web
That's everything we have on AI Safety Evaluation right now