AI Pulse by Inblix

Topic: inference-efficiency

1 article

Explore our coverage of inference-efficiency — 1 curated articles, summaries, and related resources from the Inblix archive.

JetBrains drops Mellum2, a 12B open model that runs 2x faster by only waking a fifth of its brain — Inblix summary
Research

JetBrains drops Mellum2, a 12B open model that runs 2x faster by only waking a fifth of its brain

Hugging Face Blog · Jun 1, 2026 · 2 min read

JetBrains just open-sourced Mellum2, a 12-billion-parameter Mixture-of-Experts model built for the unglamorous, high-fr...