text-to-speech
3 published articles
Open-weight speech for production voice agents
Magpie TTS spends 32ms of your voice agent's latency budget
Magpie TTS reports 32ms time-to-first-audio on an NVIDIA B200 and adds Arabic, Korean and Brazilian Portuguese, bringing its roster to 12 languages. The open-weights pitch: self-hosted speech synthesis no longer loses the latency argument.
2026-08-16
AI & Research
Voxtral TTS's 68% win rate just made ElevenLabs look vulnerable
Voxtral TTS sets a new bar for multilingual voice cloning, winning 68.4% of preference comparisons against ElevenLabs Flash v2.5. The model's hybrid architecture and open-weight release could reshape the TTS landscape.
2026-07-24
Artificial Intelligence
Nvidia's new audio model does five jobs at once and beats the specialists at their own game
Nvidia's Audex unifies audio understanding, generation, and text reasoning in a single model, matching or beating task-specific systems on speech and audio benchmarks without sacrificing text performance.
2026-07-09