Checked for new stories 10m ago

Updates on Reinforcement Learning

Every AI story we track on Reinforcement Learning — 28 stories so far, each summarized in our own words and linked back to the publisher that reported it.

Pulled from 123 sources

This month

AI Research6 min read

From MIT to IBM, expediting AI and quantum deployment

MIT AI News
Robotics3 min read

AI Sapiens K1

Hacker News
Robotics2 min read

Hugging Face Releases Programmable, Walking Duck Robot

AI Business
Robotics7 min read

Hugging Face is selling a $399 robot duck with open-source software

The Next Web

Hugging Face Unveils Microduck: A $399 Open-Source 25 cm Biped You Train with Reinforcement Learning

MarkTechPost
Machine Learning6 min read

Deep Cogito Raises $43M Series A to Build the Post-Training Engine for Self-Improving AI

Unite.AI
AI Research5 min read

Google DeepMind Extends 15 Years of Game AI Research Into EVE Online

Covered by 2 sources
AI Research5 min read

Inherent, founded by DeepMind alumni, says its AI ‘teammate’ just outperformed Anthropic and OpenAI at replicating research

TechCrunch
Military4 min read

Smack Technologies raises $61m as the Pentagon’s hurry becomes a battlefield-AI business model

The Next Web
Machine Learning4 min read

GLM-5.3 Scores 60 on Artificial Analysis Intelligence Index, Matching Kimi K3

Unite.AI
Dev5 min read

ByteDance Seed and Tsinghua AIR Introduces CUDA Agent: A Large-Scale Agentic RL System for CUDA Kernel Generation

MarkTechPost
AI Research5 min read

The Week’s 10 Biggest Funding Rounds: Data, Neolab, AI Infrastructure, Defense And AI Coding Lead

Crunchbase
Machine Learning3 min read

River AI raised $1.1bn to let companies train and keep their own models

Covered by 4 sources

5 useful things you'll learn in my new post-training textbook (shipping now!)

Interconnects
Machine Learning6 min read

Beating GPT-5.6 Sol on retrieval with 100x cheaper open models

Hacker News
AI Research1 min read

We need to RL less

LessWrong
Machine Learning1 min read

RLVR that rewards red teaming the training environment

LessWrong

Reward Laundering: LLMs Can Gain Unintended Behaviors by Deciding When to Earn Their Rewards

LessWrong
Agents25 min read

Echoverse: Deep, evolving environments for computer-use agents

Microsoft Research Blog
Machine Learning5 min read

Kimi AI and kvcache-ai Open Sources ‘AgentENV’: A Distributed System that Powers Agentic Reinforcement Learning (RL) Training for Kimi K3

MarkTechPost
AI Research1 min read

RL & search is a terrifying way to build AGI (an FAQ)

Alignment Forum
Robotics10 min read

The State of Simulation for Physical AI: An Overview

Hugging Face
Machine Learning4 min read

The father of reinforcement learning is leaving Carmack to build his own AI

The Next Web
Agents4 min read

Prime Intellect raises $130M Series A to help enterprises build their own AI agents

TechCrunch

Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing

Import AI

Overcoming reward signal challenges: Verifiable rewards-based reinforcement learning with GRPO on SageMaker AI

AWS Blog

vLLM V0 to V1: Correctness Before Corrections in RL

Hugging Face

Keep the Tokens Flowing: Lessons from 16 Open-Source RL Libraries

Hugging Face
That's everything we have on Reinforcement Learning right now