voice cloning
5 published articles
Edge AI
Audio8's CPU-only runtime fits voice cloning in about 1 GiB of RAM
Audio8's ONNX runtime runs the 0.6B TTS preview entirely on CPU: INT4 autoregressive weights, a bundled 44.1 kHz codec, streaming PCM, and an OpenAI-compatible endpoint. The service takes about 1 GiB of RAM on a laptop, and no PyTorch or Transformers are needed at runtime.
2026-08-05
Open Source TTS
Audio8's new TTS model clones voices in 11 languages for free
Audio8 releases Audio8-TTS Preview under Apache 2.0, a 0.6B model supporting 11 languages, zero-shot voice cloning, and a bundled 44.1 kHz codec. The release targets edge speech applications that need speed and privacy.
2026-07-30
AI & Research
Voxtral TTS's 68% win rate just made ElevenLabs look vulnerable
Voxtral TTS sets a new bar for multilingual voice cloning, winning 68.4% of preference comparisons against ElevenLabs Flash v2.5. The model's hybrid architecture and open-weight release could reshape the TTS landscape.
2026-07-24
Artificial Intelligence
The 'um' that makes AI sound human is finally here
MiniMax's Speech 2.8 adds native filler words, breathing, and hesitation to AI speech, solving the 'too perfect' problem that made synthetic voices robotic. The model also delivers 10-second voice cloning and cross-language accuracy.
2026-07-11
AI Speech
Minimax speech 2.8 brings human warmth to AI voices with native filler words and high-fidelity cloning
MiniMax Speech 2.8 adds natural filler words, 10-second voice cloning, and studio-quality audio to close the gap between synthetic and human speech. The model fixes the 'too perfect' problem that makes AI voices sound robotic.
2026-07-08