Today's stories Scoop: OpenAI faces GOP-led Senate investigation into Hugging Face breach A T T Covered by 3 sources 6h ago Paul Christiano joins OpenAI Foundation Board O B U +3 Covered by 6 sources 22h ago We have started losing control of AI. It’s time to shut it down | Garrison Lovely T The Guardian 4h ago OpenAI not on track to reduce risk of ‘catastrophic’ loss of control, says board member T The Guardian 2h ago Why some experts increasingly fear AI will take over B BBC 15h ago Anthropic has a cute graphic showing how its AI spread 'malicious' code B Business Insider 14h ago This week Anthropic's Alignment Science lead says there is a ">10%" chance AI could kill all humans within the next decade and is worried about recursive self-improvement (Evan Hubinger/@evanhub) T T T +6 Covered by 9 sources 1d ago Anthropic researcher Jacob Coxon says he is quitting the AI industry over fears that tech companies are racing to build systems they won't be able to control (Amrith Ramkumar/Wall Street Journal) T B T +8 Covered by 11 sources 2d ago The AI warnings are coming from inside the lab P Platformer 2d ago Fields medalist Jacob Tsimerman, set to join OpenAI later in September, launches the Mathematical AI Safety Institute to apply higher math to AI safety problems (Siobhan Roberts/New York Times) T Techmeme 2d ago “Some agents will be pursuing their own objectives”: OpenAI’s chief scientist warns AI could trick and blackmail humans T The New Stack (AI) 3d ago OpenAI admits to German wiki ‘incident’ T B T +2 Covered by 5 sources 5d ago The US plans to raise AI-directed cyberattacks with China, Nikkei reports T The Next Web 3d ago No monitoring system caught the German wiki. Two outside researchers found it by searching the internet. T The Next Web 3d ago An in-depth look at OpenAI's wiki incident: other hacked message boards, OpenAI's cover-up, how harmless web search tasks led agents to break out, and more (Zvi Mowshowitz/Don't Worry About the Vase) T Techmeme 3d ago How Rationalism, a movement pioneered by Eliezer Yudkowsky focused on existential superintelligent AI risks, influenced top AI leaders and their alarmist claims (Cal Newport/New York Times) T Techmeme 3d ago AI agents keep finding ways to bend the rules. Here are some of the wildest. B Business Insider 4d ago In “An Alien Mind,” OpenAI’s Jakub Pachocki Urges Shared Safety Bars U Unite.AI 4d ago Astra appears to think without showing its work, and the people arguing about it co-wrote the warning T The Next Web 5d ago Claude Fable 5.1 and Mythos 5.1: The System Card D Don't Worry About the Vase 6d ago An Open Letter to Bernie Sanders: Regulate AI’s Dangers, Don’t Ban Its Promise U Unite.AI 5d ago Researchers Document OpenAI Agent Swarm That Repurposed German Wiki U F A Covered by 3 sources 6d ago OpenAI’s rogue agents keep escaping, with no formal process to investigate them T TechCrunch 6d ago This month This Is the Worst Possible Time for OpenAI to BfЖ7!م#2猫$9&क G T T Covered by 3 sources 8d ago OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities W B T +1 Covered by 4 sources 9d ago TechCrunch Disrupt 2026’s new Real World AI Stage features Nvidia, robots, and extinct animals T TechCrunch 8d ago Researchers fear safety disaster ahead of OpenAI’s Astra release T The Verge 8d ago Anthropic Says It Hit the Brakes on AI Testing Following Autonomous Hacks G A D Covered by 3 sources 9d ago Anthropic Deliberately Trained an Extremely Misaligned, Reward-Seeking AI and It Did Some REALLY Bad Things F Futurism 8d ago The Singularity Is Not What It Seems: Whatever the AI Future Is, We're in It Now H Hacker News 8d ago Anthropic Announces Enterprise Frontier Safeguards, Customer-Held Data U M Covered by 2 sources 9d ago OpenAI delayed its new model’s development after the Hugging Face hack T The Verge 9d ago ‘Not perfectly aligned’ with human values: Anthropic admits security failures behind AI hacking incidents T The Guardian 9d ago Anthropic has resumed the tests in which its models attacked real companies T The Next Web 9d ago Anthropic’s Claude fixed all 10 alignment failures. Then it tried to cheat 2.4% of the time. T The New Stack (AI) 10d ago CCTV-Affiliated Account Attacks Anthropic, Sets Terms for US-China AI Talks U Unite.AI 10d ago Number of AI agents going out of control peaks in July — British newspaper T TASS English 12d ago Would you share your weirdest agent logs with AI safety researchers? H Hacker News 12d ago Sharp rise in incidents of AI escaping users’ control, research finds T The Guardian 12d ago OpenAI Says AGI Is Coming By Year-End. It Also Just Had The Worst Safety Crisis In Its History. F Forbes 13d ago This Is How Anthropic Thinks AI Agents Should Navigate the Physical World W T A Covered by 3 sources 14d ago Silico, a tool for researchers to understand their AI models better I IEEE Spectrum 15d ago Why AI Flatters You And Who Gets Paid To Stop It F Forbes 17d ago OpenAI wants California to toughen the AI law it once fought T The Next Web 17d ago Microsoft Moves AI Governance from Policy to Runtime Enforcement I InfoQ (AI, ML & Data) 17d ago Sam Altman says he's worried about AI being controlled by a few powerful players B Business Insider 18d ago OpenAI says California should strengthen its AI safety bill T E Covered by 2 sources 19d ago OpenAI pauses training of new AI models due to cyber risks — company T B H +7 Covered by 10 sources 23d ago OpenAI Halts AI Training on Advanced Model as It Detects Dark Signs Emerging F Futurism 21d ago OpenAI institutes new safeguards after Hugging Face breach T W T +3 Covered by 6 sources 23d ago OpenAI Scales Back AI Development, but it Could be Too Late A AI Business 22d ago OpenAI’s Greg Brockman: Z.ai’s GLM-5.3 likely to “significantly accelerate the threat landscape” T The New Stack (AI) 23d ago I Turned AI to the Dark Side H Hacker News 23d ago SPAR – Fall 2026 AI Safety Research Projects H Hacker News 28d ago Experts are warning: our AI arms race is putting humanity at risk | Stuart Russell T The Guardian 30d ago Bernie Sanders calls on Silicon Valley to ‘pause AI development’ in interest of humanity T B Covered by 2 sources Aug 10 Import AI 468: 23 RSI ideas; PostTrainBench+; and how trust and transparency interplay with AI racing I Import AI Aug 10 House Democrats Press Johnson for AI CEO Testimony After Rogue Model Hacks U G T Covered by 3 sources Aug 10 OpenAI's Next AI Model Astra Shows Cyber Performance Strong Enough to Trigger Pause T E Covered by 2 sources Aug 10 The AI safety test is becoming a safety risk T TechCrunch Aug 9 Showing the 60 most recent of 178 stories on AI Safety