edge AI
7 published articles
Edge AI
Audio8's CPU-only runtime fits voice cloning in about 1 GiB of RAM
Audio8's ONNX runtime runs the 0.6B TTS preview entirely on CPU: INT4 autoregressive weights, a bundled 44.1 kHz codec, streaming PCM, and an OpenAI-compatible endpoint. The service takes about 1 GiB of RAM on a laptop, and no PyTorch or Transformers are needed at runtime.
2026-08-05
Open Source TTS
Audio8's new TTS model clones voices in 11 languages for free
Audio8 releases Audio8-TTS Preview under Apache 2.0, a 0.6B model supporting 11 languages, zero-shot voice cloning, and a bundled 44.1 kHz codec. The release targets edge speech applications that need speed and privacy.
2026-07-30
Physical AI
Nvidia's 4B model does what most small robots can't: act, not just talk
Nvidia's Cosmos 3 Edge packs world modeling into 4B parameters, ranking first on VANTAGE-Bench among its size class. Its dual-transformer design blends reasoning with real-time action prediction for robots, all deployable on Jetson and RTX hardware.
2026-07-28
Edge AI
Mistral Nano runs on $80 hardware and matches 85% of its bigger sibling. That gap decides everything.
Mistral AI's new Mistral Nano targets edge devices with under 1GB of RAM, achieving 85% of the reasoning performance of its larger Mistral Small model. That 15% gap is either irrelevant or decisive, depending on what you need it to do.
2026-07-25
Enterprise AI
Mistral's enterprise pitch: own the stack, not just the model
Mistral AI unveils a comprehensive enterprise platform tailored for sales, engineering, compliance, support, and operations, with industry-specific solutions for finance, healthcare, logistics, eCommerce, government, and defense. The company highlights existing customers like BNP Paribas, AXA, CMA CGM, and France Travail.
2026-07-18
LLMs & Models
Mistral's cascade recipe shrinks LLMs without killing reasoning
Mistral's cascade distillation shrinks large models into small ones while preserving reasoning and vision. The 3B variant packs capabilities that used to require ten times the parameters, and it's all Apache 2.0.
2026-07-17
special report / edge ai
Nvidia just cracked open DeepStream. Your edge AI project will never be the same.
The full source code of Nvidia's DeepStream video analytics SDK is now on GitHub under Apache 2.0 and CC-BY-4.0, opening edge AI development to a wider audience. Version 9.1 brings LLM-based coding agents, Triton Inference Server integration, and consolidated repositories for end-to-end pipelines.
2026-07-16