OpenAI and 9 national labs throw a 1,000-scientist AI jam session
Curated by the Inblix editorial team
Picture a thousand scientists, nine national labs, and a single day to see just how much AI can bend the curve on real research. That’s what happened today as OpenAI teamed up with the Department of Energy’s sprawling lab network for a first-of-its-kind event. They called it the ‘1,000 Scientist AI Jam Session,’ and the name isn’t just branding—researchers from Argonne to Princeton Plasma Physics spent the day stress-testing frontier models like o3-mini on their own thorny scientific problems. The point wasn’t just to see if the AI was smart enough. It was to figure out if it was useful enough, in the hands of people who know the difference.
The guest list tells you this is a priority. U.S. Secretary of Energy Chris Wright and OpenAI President Greg Brockman showed up at Oak Ridge National Laboratory, touring the floor and talking with scientists in real time. Wright didn’t mince words, framing the effort in starkly nationalistic terms. ‘Like the Manhattan Project, which brought together the world’s best scientists and engineers for a patriotic effort that changed the world, AI development is a race that the United States must win,’ he said. That’s not a subtle metaphor. It sets a clear expectation: this collaboration isn’t a science fair, it’s a front in a broader competition.
Beyond the one-day event, this is a down payment on a longer relationship. Just last month, OpenAI inked a deal to deploy an o-series reasoning model across the labs, targeting breakthroughs in materials science, renewable energy, and astrophysics. There’s also a parallel track with Los Alamos focused on safely using multimodal models for bioscientific research. The common thread is turning the labs’ staggering data reserves into insights that actually leave the lab. The participating institutions—Livermore, Brookhaven, Berkeley, and others—represent a massive concentration of computing power and domain expertise that most AI companies can only dream of accessing.
The follow-up will be the real test. OpenAI and the labs plan to publish a report on their findings, essentially a public report card on where frontier models actually help working scientists versus where they still stumble. That kind of candid feedback loop is rare in the AI world, which tends to prefer glossy demos over honest assessments. If the report delivers specifics—exactly which problems o3-mini cracked and which ones made it hallucinate—it could shape how the next generation of scientific AI gets built. For now, a thousand researchers just got a day to kick the tires. The rest of us will have to wait to see what they found.
💡 Key Takeaways
- OpenAI's o3-mini model is being tested by actual domain experts across nine national labs, moving beyond benchmark scores to real-world scientific utility.
- Secretary Wright explicitly compared the AI push to the Manhattan Project, signaling that the administration views public-private AI partnerships as a matter of national security and competitiveness.
- The event follows a recent, concrete deal to deploy OpenAI's reasoning models at the labs, with specific target areas including materials science, renewable energy, and biosafety research.
- A promised follow-up report on the scientists' findings could provide an unusually candid, large-scale evaluation of where frontier AI models genuinely accelerate research and where they fall short.
Keep reading: See related articles below for more coverage on this topic.
Get smarter about AI
The sharpest AI news, curated daily. Delivered free to your inbox.