How remembering failure made robots learn 8x faster
OpenAI Blog · Jul 20, 2026 · 3 min read
Here's a problem anyone who's trained a robot knows too well: most of the time, the robot fails. And in reinforcement l...
Browse 1135 curated Product articles — AI news summaries covering OpenAI, reinforcement-learning, robotics, and more.
1135 articles
OpenAI Blog · Jul 20, 2026 · 3 min read
Here's a problem anyone who's trained a robot knows too well: most of the time, the robot fails. And in reinforcement l...
OpenAI Blog · Jul 20, 2026 · 3 min read
OpenAI just anointed Proximal Policy Optimization as its go-to reinforcement learning algorithm, and for anyone who's w...
OpenAI Blog · Jul 20, 2026 · 2 min read
Remember last week's reassuring claim—that self-driving cars couldn't easily be tricked by altered street signs because...
OpenAI Blog · Jul 20, 2026 · 2 min read
Throwing carefully calibrated noise at a neural network's weights, rather than its actions, can dramatically speed up h...
OpenAI Blog · Jul 20, 2026 · 2 min read
OpenAI just open-sourced RL-Teacher, a lean system that lets you train reinforcement learning agents using direct human...
OpenAI Blog · Jul 20, 2026 · 3 min read
The timeline is almost absurd when you lay it out chronologically. In early May 2017, OpenAI's Dota 2 1v1 bot was losin...
OpenAI Blog · Jul 20, 2026 · 2 min read
An AI bot built by OpenAI has done something quietly terrifying. It didn't just beat a pro at Dota 2 — it dismantled on...
OpenAI Blog · Jul 20, 2026 · 2 min read
The core headache of multi-agent AI isn't just getting agents to learn — it's getting them to learn while everyone else...
OpenAI Blog · Jul 20, 2026 · 3 min read
The assumption is so foundational it’s practically dogma: stack a bunch of linear layers without a nonlinearity like Re...
OpenAI Blog · Jul 20, 2026 · 2 min read
Reinforcement learning agents are brilliant at many things, but they've always had a glaring blind spot: other agents....
OpenAI Blog · Jul 20, 2026 · 2 min read
Here's a research finding that actually feels like a scene out of a robot boxing movie. A team has shown that by tweaki...
OpenAI Blog · Jul 20, 2026 · 2 min read
OpenAI has demonstrated that simulated humanoid robots can develop complex physical strategies like tackling, ducking,...
OpenAI Blog · Jul 20, 2026 · 2 min read
Here’s a finding that flips a core assumption in robotics on its head: you don’t need realistic training data to teach...
OpenAI Blog · Jul 20, 2026 · 2 min read
Getting a robot policy to work in a simulation and then actually function in the real world remains one of the hardest...
OpenAI Blog · Jul 20, 2026 · 2 min read
The oldest problem in robotics might have just found its simplest solution. For decades, engineers have grappled with t...
OpenAI Blog · Jul 20, 2026 · 2 min read
OpenAI has cracked a persistent robotics problem: getting simulated training to actually work on physical hardware. The...
OpenAI Blog · Jul 20, 2026 · 2 min read
Reinforcement learning has a scaling problem. When a task demands thousands of individual steps, brute-forcing a soluti...
OpenAI Blog · Jul 20, 2026 · 2 min read
Machine learning researchers have long known that neural networks make terrible teachers — not because they fail to edu...
OpenAI Blog · Jul 20, 2026 · 2 min read
Here's a problem that's annoyed researchers for years: the L₀ norm — essentially counting how many non-zero weights a m...
OpenAI Blog · Jul 20, 2026 · 2 min read
The bottleneck in deep learning isn't just ideas — it's what the hardware can run efficiently. For years, sparse neural...
OpenAI Blog · Jul 20, 2026 · 3 min read
Entity disambiguation—the task of figuring out whether 'jaguar' means the car, the animal, or the sports team—has a new...
OpenAI Blog · Jul 20, 2026 · 2 min read
Running a Kubernetes cluster at 500 nodes is one thing. Pushing past 2,500 is a parade of cascading failures that the O...
OpenAI Blog · Jul 20, 2026 · 2 min read
OpenAI just published a fresh set of seven unsolved research problems, a spiritual successor to their original 'Request...
OpenAI Blog · Jul 20, 2026 · 3 min read
A new multi-institutional report is drawing a bullseye on a future where artificial intelligence doesn't just malfuncti...