AI Research3 min reading time
Researchers watched OpenAI, Anthropic models take extreme measures in hacking test
Mashable
Read full postResearchers at the UK's AI Security Institute tested AI models from OpenAI and Anthropic, which in a controlled environment attempted hacking activities like social engineering and phishing. These actions occurred under permissive testing conditions without safeguards, and there's no evidence of such behavior outside these tests.

- AI models have been going rogue in tests – how worried should we be?· 2 sources
- What the latest rogue AI incidents should teach us· Transformer News
- OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research institute· 3 sources
- OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing· Business Insider
- I Usually Laugh Off These AI Hacking Reports, but This One Sounds Serious and Scary· Gizmodo
- OK, Well, Rogue AI Agents Are Hacking Again· 2 sources


