
Checked for new stories 19m ago
Updates on Reinforcement Learning
Every AI story we track on Reinforcement Learning — 29 stories so far, each summarized in our own words and linked back to the publisher that reported it.
Pulled from 124 sources
This month




Sources: OpenAI bought tens of thousands of Macs for RL, Anthropic rents them, Nvidia sees Apple as its main local AI rival as Macs gain traction with AI devs (Aaron Tilley/The Information)
Covered by 2 sources

Hugging Face Unveils Microduck: A $399 Open-Source 25 cm Biped You Train with Reinforcement Learning
MarkTechPost

Machine Learning6 min read
Deep Cogito Raises $43M Series A to Build the Post-Training Engine for Self-Improving AI
Unite.AI

AI Research5 min read
Google DeepMind Extends 15 Years of Game AI Research Into EVE Online
Covered by 2 sources

AI Research5 min read
Inherent, founded by DeepMind alumni, says its AI ‘teammate’ just outperformed Anthropic and OpenAI at replicating research
TechCrunch

Military4 min read
Smack Technologies raises $61m as the Pentagon’s hurry becomes a battlefield-AI business model
Covered by 2 sources

Machine Learning4 min read
GLM-5.3 Scores 60 on Artificial Analysis Intelligence Index, Matching Kimi K3
Unite.AI

Dev5 min read
ByteDance Seed and Tsinghua AIR Introduces CUDA Agent: A Large-Scale Agentic RL System for CUDA Kernel Generation
MarkTechPost

AI Research5 min read
The Week’s 10 Biggest Funding Rounds: Data, Neolab, AI Infrastructure, Defense And AI Coding Lead
Crunchbase

Machine Learning3 min read
River AI raised $1.1bn to let companies train and keep their own models
Covered by 4 sources

LLM & Text Generation8 min read
5 useful things you'll learn in my new post-training textbook (shipping now!)
Interconnects



LLM & Text Generation1 min read
Reward Laundering: LLMs Can Gain Unintended Behaviors by Deciding When to Earn Their Rewards
LessWrong

Agents25 min read
Echoverse: Deep, evolving environments for computer-use agents
Microsoft Research Blog

Machine Learning5 min read
Kimi AI and kvcache-ai Open Sources ‘AgentENV’: A Distributed System that Powers Agentic Reinforcement Learning (RL) Training for Kimi K3
MarkTechPost


Machine Learning4 min read
The father of reinforcement learning is leaving Carmack to build his own AI
The Next Web

Agents4 min read
Prime Intellect raises $130M Series A to help enterprises build their own AI agents
TechCrunch

Import AI 460: Reward hacking society, RSI data from Anthropic; and RL-based quadcopter racing
Import AI

Overcoming reward signal challenges: Verifiable rewards-based reinforcement learning with GRPO on SageMaker AI
AWS Blog
That's everything we have on Reinforcement Learning right now
