speech synthesis
4 published articles
Speech synthesis
Alibaba's new TTS model speaks 16 languages, laughs on command, and might actually ship
Alibaba's Qwen-Audio-3.0-TTS is a production-oriented speech synthesis system that combines a low-frame-rate tokenizer with progressive training. It supports 16 languages, 20 Chinese dialect regions, and natural-language instructions for emotion, pace, and speaking style.
2026-07-26
Artificial Intelligence
The 'um' that makes AI sound human is finally here
MiniMax's Speech 2.8 adds native filler words, breathing, and hesitation to AI speech, solving the 'too perfect' problem that made synthetic voices robotic. The model also delivers 10-second voice cloning and cross-language accuracy.
2026-07-11
AI Speech
Minimax speech 2.8 brings human warmth to AI voices with native filler words and high-fidelity cloning
MiniMax Speech 2.8 adds natural filler words, 10-second voice cloning, and studio-quality audio to close the gap between synthetic and human speech. The model fixes the 'too perfect' problem that makes AI voices sound robotic.
2026-07-08
Artificial Intelligence
MiniMax just shipped a model for every AI job you can name
Chinese AI startup MiniMax launches M3, Hailuo 2.3, MiniMax Code, and new speech/music models, broadening its product lineup in a competitive landscape.
2026-07-06