SevenTnewS

open source

75 published articles

Open Source4 min read

Open Source · Vibe Design Workspace

Open Design takes on Claude Design from your own machine

Open Design, an Apache-2.0 vibe design workspace, positions itself as the open-source local answer to Claude Design. Your agent, your keys, your files: no vendor cloud, no subscription, 21 coding agents and 85.5K+ GitHub stars.

2026-08-22

Qwen / Alibaba5 min read

AI Safety: Abliteration and Open Weights

Abliterated Qwen3.8-27B: refusals drop to 0%, benchmarks barely move

An abliterated, FP8-quantized build of Qwen3.8-27B refuses 0% of harmful prompts on AdvBench, down from 99%, while general benchmarks stay within 1.3 points. The model card documents the method in unusual detail. The caveats deserve equal attention.

2026-08-16

DeepSeek4 min read

AI Agents · Open Source

DeepSeek ships an agent harness where even the model is a plugin

DeepSeek released DeepSeek Harness, an open-source agent runtime where models, tools, sandboxes, and the UI are all Cordis plugins. Append-only session logs and a two-tool minimal mode point to a quieter ambition: auditable, reproducible agent runs.

2026-08-16

AI4 min read

Multimodal / Open Source

Qwen plugin pack gives your coding agent eyes, hands, and video memory

Alibaba's Qwen team released Qwen-MM-Plugins, an Apache-2.0 toolkit that gives coding agents native multimodal skills: dynamic-resolution image and video reading, OCR, grounding, ASR, long-video memory, video generation, and thin-client control of Blender and FreeCAD. One script installs it across Claude Code, Codex, Qoder, OpenClaw, Qwen Code, and Gemini CLI.

2026-08-15

Meta AI3 min read

Open Source AI

Meta's 30B Muse Glimmer hits Apple Silicon first, NVIDIA support follows

Meta released Muse Glimmer, a 30B open-weights multimodal model built for agent workloads, and Ollama shipped it for Apple Silicon the same day. DFlash makes it 1.5x to 1.8x faster on Mac hardware. NVIDIA and AMD support follow in the coming days.

2026-08-13

Meta AIFeatured4 min read

Open source AI

Muse Glimmer: Meta's 30B agent fits under 20GB, cloud optional

Meta open-sourced Muse Glimmer, a 30B agentic model that runs offline on a single consumer GPU. Quantization keeps it under 20 GB; a DFlash drafter delivers up to 3.1x faster decoding on an RTX 5090, and the weights are on Hugging Face under Apache 2.0.

2026-08-10

Qwen / Alibaba4 min read

Open Source AI

Alibaba's biggest model coded alone for 16 days. Next week, its weights go public.

Qwen3.8-Max activates just 95 billion of its 2.4 trillion parameters, ranks second in Vision Arena, and built an agent framework in a 16-day autonomous run. Alibaba publishes the weights next week in its first open-weight flagship release.

2026-08-10

Qwen / Alibaba4 min read

Qwen / Alibaba Cloud

Qwen2.5-Omni outperforms Gemini-1.5-Pro on OmniBench, fits under 12GB

Alibaba's open-source Qwen2.5-Omni outscored Gemini-1.5-Pro on OmniBench and topped the MMAU audio reasoning leaderboard. Quantized builds cut VRAM below 12GB and MNN support brings real-time voice chat to phones.

2026-08-07

AI4 min read

AI Video Generation

MiniMax H3 prices video at a third of rivals, ranks No. 1 in editing

MiniMax H3 packs text, image, video and audio into one model, prices output at a third of mainstream per-second rates and ranks first in video editing on Artificial Analysis. The weights are promised within days. How the output holds up outside the company's own demos is the open question.

2026-08-06

Qwen / Alibaba4 min read

Qwen3.8-Max open weights land next week

Alibaba's most powerful model ever is going open source

Qwen3.8-Max, Alibaba's first open-weight Max-class model at 2.4 trillion parameters, hits Hugging Face and ModelScope next week. The launch reframes the open-source question: what happens when the frontier's biggest weights are free to download and test?

2026-08-06

AI4 min read

Edge AI

Audio8's CPU-only runtime fits voice cloning in about 1 GiB of RAM

Audio8's ONNX runtime runs the 0.6B TTS preview entirely on CPU: INT4 autoregressive weights, a bundled 44.1 kHz codec, streaming PCM, and an OpenAI-compatible endpoint. The service takes about 1 GiB of RAM on a laptop, and no PyTorch or Transformers are needed at runtime.

2026-08-05

Open Source6 min read

Open Source

A relay race of rented GPUs trained NanoColibri's 2.7B MoE for $200

NanoColibri-Instruct went from blank weights to a working 2.7B MoE for about $200. Volunteers passed a training baton on the Hugging Face Hub, one rented GPU at a time, with a compare-and-swap lease so no two people ever trained the same leg.

2026-08-04

← PreviousPage 1 / 7 · 75 articlesNext →