AI Pulse by Inblix

GPT-5.3-Codex-Spark: Real-time AI coding preview

OpenAI Blog · Jul 11, 2026 · 1 min read · Read original article →

Curated by the Inblix editorial team


OpenAI just dropped a research preview of GPT-5.3-Codex-Spark, a smaller, faster version of their code model built for real-time collaboration. It’s the first fruit of their partnership with Cerebras, delivering over 1000 tokens per second on ultra-low latency hardware. Think of it as the code companion that keeps up with your pace—you can interrupt it mid-thought, redirect it, or ask for tweaks and see results instantly. While it’s text-only with a 128k context window for now, it’s a taste of a bigger vision: combining long-running autonomous coding sessions with instant, interactive editing. ChatGPT Pro users get first dibs to experiment, though you might hit rate limits or queues as they balance demand. Why it matters: Codex-Spark signals a shift toward AI that works with you, not just for you—where response time is as crucial as raw intelligence, making real-time pair programming with AI feel like a natural extension of the developer’s workflow.

💡 Key Takeaways

  1. Codex-Spark is a smaller version of GPT-5.3-Codex optimized for real-time, low-latency coding with over 1000 tokens per second.
  2. It's a research preview exclusive to ChatGPT Pro users, with a 128k context window and text-only input, but features like automatic test execution are disabled by default.
  3. Performance on coding benchmarks like SWE-Bench Pro shows strong results in a fraction of the time compared to the larger model.

Keep reading: See related articles below for more coverage on this topic.

Get smarter about AI

The sharpest AI news, curated daily. Delivered free to your inbox.

← Back to all articles