Why counting states still beats complex exploration in deep RL
OpenAI Blog · Jul 20, 2026 · 2 min read
For years, the consensus was clear: count-based exploration works beautifully in small, tabular settings but falls apar...
7 articles
Explore our coverage of exploration — 7 curated articles, summaries, and related resources from the Inblix archive.
OpenAI Blog · Jul 20, 2026 · 2 min read
For years, the consensus was clear: count-based exploration works beautifully in small, tabular settings but falls apar...
OpenAI Blog · Jul 20, 2026 · 2 min read
That mysterious Q* project just got a little less mysterious—and a lot more interesting for anyone who cares about maki...
OpenAI Blog · Jul 20, 2026 · 2 min read
The hardest problems in reinforcement learning aren't about mastering a known skill — they're about figuring out what t...
OpenAI Blog · Jul 19, 2026 · 2 min read
OpenAI just posted a staggering 74,500 points on the notoriously difficult Atari game Montezuma's Revenge, beating any...
OpenAI Blog · Jul 19, 2026 · 2 min read
Reinforcement learning has an obsession with rewards. For years, the playbook has been the same: carefully design a den...
OpenAI Blog · Jul 19, 2026 · 2 min read
There's a reason reinforcement learning researchers have been obsessed with a dusty 1984 Atari game. Montezuma's Reveng...
OpenAI Blog · Jul 19, 2026 · 2 min read
A new framework called POLO — plan online, learn offline — shows how simulated robots can master genuinely difficult ph...