NVIDIA's 11B voice model kills the ASR-LLM-TTS chain — and brings tool calling
MarkTechPost · Aug 9, 2026 · 3 min read
NVIDIA dropped an 11-billion-parameter model that finally collapses the clunky ASR-to-LLM-to-TTS pipeline into a single...
8 articles
Explore our coverage of speech-to-speech — 8 curated articles, summaries, and related resources from the Inblix archive.
MarkTechPost · Aug 9, 2026 · 3 min read
NVIDIA dropped an 11-billion-parameter model that finally collapses the clunky ASR-to-LLM-to-TTS pipeline into a single...
OpenAI Blog · Jul 15, 2026 · 2 min read
The era of stringing together three separate models just to build a voice assistant is officially over. OpenAI moved it...
TechCrunch AI · Jul 15, 2026 · 2 min read
The enterprise voice AI space is getting crowded, but San Francisco-based Rime just landed $24 million in Series A fund...
Hugging Face Blog · Jul 15, 2026 · 2 min read
We've all been there. You ask a voice assistant something, it transcribes every word perfectly, but you can just tell i...
Hugging Face Blog · May 27, 2026 · 2 min read
A new tutorial from Hugging Face shows how to run Reachy Mini, the open-source robot, completely offline. No cloud. No...
Hugging Face Blog · Mar 24, 2026 · 2 min read
The industry has been grading voice agents all wrong. Accuracy and conversational experience aren't separate report car...
Hugging Face Blog · Dec 20, 2024 · 2 min read
Give a state-of-the-art model a logic puzzle in writing, and it aces it. Read the same puzzle aloud, and that performan...
Hugging Face Blog · Oct 22, 2024 · 2 min read
Hugging Face just dropped an open-source Speech-to-Speech pipeline that feels a bit like magic: you talk, the machine t...