Product
OpenAI's algorithm learns backflip from under 900 bits of human feedback
OpenAI Blog · Jul 20, 2026 · 2 min read
Writing reward functions that perfectly capture complex human goals is a notoriously hard problem, and getting it even...
2 articles
Explore our coverage of human-feedback — 2 curated articles, summaries, and related resources from the Inblix archive.
OpenAI Blog · Jul 20, 2026 · 2 min read
Writing reward functions that perfectly capture complex human goals is a notoriously hard problem, and getting it even...
OpenAI Blog · Jul 19, 2026 · 2 min read
Training AI with human feedback sounds like a straightforward path to better, safer models. But new research from OpenA...