AI Pulse by Inblix

OpenAI forms 'Preparedness' team to fight AI's darkest risks

OpenAI Blog · Jul 17, 2026 · 2 min read · Read original article →

Curated by the Inblix editorial team


Featured image for article: OpenAI forms 'Preparedness' team to fight AI's darkest risks

OpenAI isn’t just building smarter models. It’s now publicly wrestling with how those models could genuinely break things — or break the world. The company detailed a new, dedicated ‘Preparedness’ team led by Aleksander Madry, an MIT professor and noted AI security researcher. The group’s mandate is refreshingly blunt: track, evaluate, forecast, and protect against catastrophic risks from frontier AI systems. We’re not talking about biased outputs here. The team’s remit spans individualized persuasion at scale, full-on cybersecurity offensives, chemical and biological threat creation, and even autonomous replication — an AI that can copy and improve itself without human input.

The announcement arrives with a concrete artifact: a Risk-Informed Development Policy (RDP). This document is meant to be the operational backbone, laying out how OpenAI will create rigorous evaluations for dangerous capabilities before a model is released, define a ladder of protective actions, and establish a governance structure with real accountability. Think of it as a safety case that evolves alongside the technology, not a one-time white paper. It’s designed to complement existing alignment work, covering the life cycle of a system both before and after it hits the market.

But the most revealing piece might be the results of the AI Preparedness Challenge. OpenAI dangled $25,000 in API credits each for the ten best ideas on catastrophic misuse, and the winning submissions paint a chillingly creative picture of what keeps researchers up at night. One winner, Claudia Biancotti, proposed using AI to deliberately precipitate a financial crisis in a strategically important country. Another, Joel Hypolite, explored causing plane crashes by hijacking radio frequencies to disrupt flight paths. Other winning concepts included scaling blackmail scams, reverse-engineering classified information, and impeding medical care access. These aren’t hypothetical sci-fi plots; they are specific, technically-grounded scenarios that the company now considers plausible enough to build defenses against.

Aleksander Madry’s appointment signals a hard turn toward empirical security research over philosophical debate. His background is in making machine learning models robust to adversarial attacks — someone who thinks like an attacker. The team is actively recruiting, and the challenge winners are on their radar as potential hires. The underlying message is clear: the conversation about AI safety is shifting from measuring alignment to actively red-teaming for catastrophic outcomes. The question isn’t just whether an AI can be helpful, but whether the safeguards can hold against someone actively trying to wield it as a weapon.

💡 Key Takeaways

  1. OpenAI's new Preparedness team, led by adversarial ML expert Aleksander Madry, will focus specifically on catastrophic risks like biological threats and autonomous replication, not just bias or safety.
  2. The AI Preparedness Challenge surfaced concrete, technically-grounded misuse scenarios, including deliberately crashing planes and triggering financial crises, which the company now uses to guide its defenses.
  3. A formal Risk-Informed Development Policy (RDP) commits OpenAI to rigorous pre-release evaluations and a governance structure for accountability, moving beyond voluntary pledges to an operational framework.

Keep reading: See related articles below for more coverage on this topic.

Get smarter about AI

The sharpest AI news, curated daily. Delivered free to your inbox.

Learn more

Glossary terms

← Back to all articles