OpenAI's New Chip Aims to Speed Up AI Inference
Curated by the Inblix editorial team
OpenAI has unveiled ‘Jalapeño,’ a custom chip designed specifically for large language model inference. The chip, developed with Broadcom, promises better performance per watt and faster running times. This move marks OpenAI’s entry into custom hardware, which could make running AI models cheaper and more reliable. Why it matters: By controlling the full stack from chip to product, OpenAI can potentially reduce costs and increase the reliability of its AI models, a significant development in the pursuit of more widespread AI adoption.
💡 Key Takeaways
- OpenAI and Broadcom have developed a custom chip called 'Jalapeño' for large language model inference
- The chip is designed to deliver better performance per watt and faster running times
- OpenAI plans to deploy the chip at scale by late 2026, with Microsoft expected to buy 40% of the chips
Keep reading: See related articles below for more coverage on this topic.
Get smarter about AI
The sharpest AI news, curated daily. Delivered free to your inbox.