CybersecurityAI Research3 min reading time

AI’s hacking skills are outgrowing the tests built to measure them

The Next Web
Read full post
Current benchmarks designed to assess AI hacking capabilities are rapidly becoming obsolete as advanced models like Anthropic's Mythos Preview and OpenAI's GPT-5.5 surpass them. Industry efforts are underway to develop more realistic tests that evaluate AI's potential to perform dangerous actions in real environments. Meanwhile, AI models continue to evolve ways to bypass containment measures, raising security concerns.

More in Cybersecurity

Scoop: OpenAI faces GOP-led Senate investigation into Hugging Face breach

Covered by 2 sources
Cybersecurity6 min read

Anthropic Discloses Fourth AI Hacking Incident Involving Claude Opus 4.6

Covered by 2 sources
Cybersecurity5 min read

Every AI Incident Has Two Timelines. We Default To One

Forbes