SevenTnewS

open weights

9 published articles

LLMs & Models4 min read

Open-weight speech for production voice agents

Magpie TTS spends 32ms of your voice agent's latency budget

Magpie TTS reports 32ms time-to-first-audio on an NVIDIA B200 and adds Arabic, Korean and Brazilian Portuguese, bringing its roster to 12 languages. The open-weights pitch: self-hosted speech synthesis no longer loses the latency argument.

2026-08-16

Meta AIFeatured3 min read

Open Source

Meta's 30B Muse Glimmer lands on Apple Silicon today via Ollama's MLX engine

Meta opened the weights for Muse Glimmer, a 30B dense model, and Ollama ships it the same day on Apple Silicon. Local coding agents gain a native backend, with Muse Spark 1.2 teased for later.

2026-08-10

LLMs & Models5 min read

Video generation

MiniMax scrapped its proven architecture to make H3 do everything

MiniMax says H3 unifies text, image, video and audio generation in one model, prices 2K output below a third of mainstream models, and plans to open the weights within days. The small print is the real story: the company abandoned the architecture that gave it an edge to get there.

2026-08-07

Qwen / Alibaba4 min read

Qwen3.8-Max open weights land next week

Alibaba's most powerful model ever is going open source

Qwen3.8-Max, Alibaba's first open-weight Max-class model at 2.4 trillion parameters, hits Hugging Face and ModelScope next week. The launch reframes the open-source question: what happens when the frontier's biggest weights are free to download and test?

2026-08-06

Mistral AIFeatured4 min read

AI Safety: 3B Classifier, Apache 2.0 Weights

Mistral's Shieldstral puts your moderation policy in the prompt, not the weights

Shieldstral frames moderation as a binary question: an instruction, a yes/no query, and the content to judge. Mistral says the 3B model matches open guardrails up to seven times its size on text safety, with Apache 2.0 weights that run on one 16GB GPU.

2026-08-05

AIFeatured5 min read

Coalition of rivals

OpenAI, Meta, and 40 others sign an AI policy letter. Here's why that matters.

Forty-one organizations from across the AI ecosystem jointly signed an open letter arguing that open-weight models are essential to American AI leadership. The unusual coalition includes rivals like OpenAI and Meta, cloud giants, and startups, signaling rare policy consensus.

2026-07-26

Mistral AI3 min read

AI Agents

Mistral coding agents leave your laptop behind: parallel sessions, no hovering required

Mistral is moving coding agents to the cloud, enabling parallel, asynchronous task execution via the new Mistral Medium 3.5 model. The update introduces remote agents in Mistral Vibe and Le Chat, along with a Work mode for multi-step tasks, while keeping human oversight for sensitive actions.

2026-07-12

LLMs & ModelsFeatured4 min read

Open Weights Analysis

Gemma 4 is infrastructure, not a chatbot. That changes the math on self-hosting AI.

Google DeepMind's Gemma 4 is an open-weight model designed for self-hosting and customization, not consumer chat. This analysis compares it to ChatGPT, Claude, and Qwen-3.5 across licensing, privacy, and deployment flexibility, revealing why it matters for regulated industries and on-device AI.

2026-07-10

DeepSeek4 min read

Open-source AI research

Million-token context on a tenth of the KV cache: DeepSeek-V4's efficiency bet

DeepSeek's V4 preview pairs a 1.6T-parameter Pro model with a 284B Flash variant, both at one million tokens of context. The paper claims 27% of the inference FLOPs and 10% of the KV cache of V3.2, the numbers that make the context cost-effective to serve.

2026-06-22