AI ResearchCybersecurity6 min reading time

A New Trick Reveals AI Models’ Inner Thoughts

Covered by 2 sources
Read full post
Researchers from University of Tübingen and partners developed a method to extract hidden reasoning steps from AI models, revealing similarities suggesting some Chinese models may have copied US models' reasoning. The method also exposed a vulnerability allowing extraction of personal data from models, which has since been fixed.

Covered by 2 sources

More on this story


More in AI Research

AI Research3 min read

Worried Anthropic researchers warn that AI ‘could kill all humans’

Covered by 8 sources
AI Research4 min read

Suno trained its v6 AI music models with help from Warner and BMG

Covered by 5 sources
AI Research4 min read

An Anthropic researcher just quit, saying OpenAI and Anthropic are 'gambling with our lives'

Covered by 10 sources