TTS
3 published articles
Edge AI
Audio8's CPU-only runtime fits voice cloning in about 1 GiB of RAM
Audio8's ONNX runtime runs the 0.6B TTS preview entirely on CPU: INT4 autoregressive weights, a bundled 44.1 kHz codec, streaming PCM, and an OpenAI-compatible endpoint. The service takes about 1 GiB of RAM on a laptop, and no PyTorch or Transformers are needed at runtime.
2026-08-05
Open Source TTS
Audio8's new TTS model clones voices in 11 languages for free
Audio8 releases Audio8-TTS Preview under Apache 2.0, a 0.6B model supporting 11 languages, zero-shot voice cloning, and a bundled 44.1 kHz codec. The release targets edge speech applications that need speed and privacy.
2026-07-30
Speech AI
Voice AI said 'I understand your frustration.' It had no idea what that meant.
Hume's Real World VoiceEQ benchmark, based on more than one million human ratings, tests over 40 voice models across dimensions standard benchmarks ignore: emotion, speaker identity, and acoustic context. The findings show speech-to-speech models vary wildly, and even leading systems often ignore the audio cues humans use instinctively.
2026-07-20