open weights
9 published articles
Open-weight speech for production voice agents
Magpie TTS spends 32ms of your voice agent's latency budget
Magpie TTS reports 32ms time-to-first-audio on an NVIDIA B200 and adds Arabic, Korean and Brazilian Portuguese, bringing its roster to 12 languages. The open-weights pitch: self-hosted speech synthesis no longer loses the latency argument.
2026-08-16
Open Source
Meta's 30B Muse Glimmer lands on Apple Silicon today via Ollama's MLX engine
Meta opened the weights for Muse Glimmer, a 30B dense model, and Ollama ships it the same day on Apple Silicon. Local coding agents gain a native backend, with Muse Spark 1.2 teased for later.
2026-08-10
Video generation
MiniMax scrapped its proven architecture to make H3 do everything
MiniMax says H3 unifies text, image, video and audio generation in one model, prices 2K output below a third of mainstream models, and plans to open the weights within days. The small print is the real story: the company abandoned the architecture that gave it an edge to get there.
2026-08-07
Qwen3.8-Max open weights land next week
Alibaba's most powerful model ever is going open source
Qwen3.8-Max, Alibaba's first open-weight Max-class model at 2.4 trillion parameters, hits Hugging Face and ModelScope next week. The launch reframes the open-source question: what happens when the frontier's biggest weights are free to download and test?
2026-08-06
AI Safety: 3B Classifier, Apache 2.0 Weights
Mistral's Shieldstral puts your moderation policy in the prompt, not the weights
Shieldstral frames moderation as a binary question: an instruction, a yes/no query, and the content to judge. Mistral says the 3B model matches open guardrails up to seven times its size on text safety, with Apache 2.0 weights that run on one 16GB GPU.
2026-08-05
Coalition of rivals
OpenAI, Meta, and 40 others sign an AI policy letter. Here's why that matters.
Forty-one organizations from across the AI ecosystem jointly signed an open letter arguing that open-weight models are essential to American AI leadership. The unusual coalition includes rivals like OpenAI and Meta, cloud giants, and startups, signaling rare policy consensus.
2026-07-26
AI Agents
Mistral coding agents leave your laptop behind: parallel sessions, no hovering required
Mistral is moving coding agents to the cloud, enabling parallel, asynchronous task execution via the new Mistral Medium 3.5 model. The update introduces remote agents in Mistral Vibe and Le Chat, along with a Work mode for multi-step tasks, while keeping human oversight for sensitive actions.
2026-07-12
Open Weights Analysis
Gemma 4 is infrastructure, not a chatbot. That changes the math on self-hosting AI.
Google DeepMind's Gemma 4 is an open-weight model designed for self-hosting and customization, not consumer chat. This analysis compares it to ChatGPT, Claude, and Qwen-3.5 across licensing, privacy, and deployment flexibility, revealing why it matters for regulated industries and on-device AI.
2026-07-10
Open-source AI research
Million-token context on a tenth of the KV cache: DeepSeek-V4's efficiency bet
DeepSeek's V4 preview pairs a 1.6T-parameter Pro model with a 284B Flash variant, both at one million tokens of context. The paper claims 27% of the inference FLOPs and 10% of the KV cache of V3.2, the numbers that make the context cost-effective to serve.
2026-06-22