SevenTnewS

Qwen

22 published articles

LLMs & Models4 min read

AI Music Generation: Qwen's answer to prompt drift

Music 3.0 swaps the one-tag prompt for a timeline that keeps AI songs on track

AI tracks tend to drift from the prompt as they unfold: instruments drop out, emotion flattens, the vocal style comes and goes. Music 3.0 swaps the one-tag description for a time-sequential Structured Caption, backed by an 8B/0.6B Hybrid-LM that splits structure from detail.

2026-08-21

AI4 min read

Multimodal / Open Source

Qwen plugin pack gives your coding agent eyes, hands, and video memory

Alibaba's Qwen team released Qwen-MM-Plugins, an Apache-2.0 toolkit that gives coding agents native multimodal skills: dynamic-resolution image and video reading, OCR, grounding, ASR, long-video memory, video generation, and thin-client control of Blender and FreeCAD. One script installs it across Claude Code, Codex, Qoder, OpenClaw, Qwen Code, and Gemini CLI.

2026-08-15

Qwen / Alibaba4 min read

AI Economics

Alibaba's $18 plan runs Qwen and DeepSeek inside Claude Code

Alibaba Cloud's Token Plan Individual bundles Qwen and third-party models such as DeepSeek into a single credit pool, from $6 a month, usable inside Claude Code and Cursor. The company claims roughly 40% savings over pay-as-you-go, but the plan only works from Singapore and only inside approved tools.

2026-08-14

AI6 min read

Cloud & LLM Security

Alibaba Cloud's AI gateway blocks prompt attacks with a 200, not a 403

A hands-on walkthrough of Alibaba Cloud's AI Gateway shows how authentication, guardrails and PII masking work as one pipeline. The gotchas: blocks return HTTP 200, key enforcement needs a manual toggle, and restored data can end up in logs.

2026-08-09

AI3 min read

AI MUSIC

The melody-first trick that lets Qwen-Music beat Suno on 13 metrics

Qwen-Music separates melody planning from full-song generation via Melody-CoT. The 3-billion-parameter model, trained on 5 million+ hours of multilingual data, achieves state-of-the-art results against Suno V5 and MiniMax Music 2.6 in both objective metrics and human preference.

2026-07-31

AI Agents3 min read

GUI Agents

Alibaba's Qwen-UI-Agent scores higher on real phones than in sandboxes

Alibaba Tongyi Lab's Qwen-UI-Agent claims state-of-the-art mobile scores (97.5% AndroidDaily) and competitive computer-use results against Opus 4.8, Gemini 3.1 Pro, and GPT-5.6 Sol. Its real-device benchmark score beats its sandbox score, while 40% partial progress on OSWorld-v2 is the honest limit.

2026-07-31

Qwen / Alibaba2 min read

Qwen / Alibaba

Alibaba splits its flagship AI model in two because one size fits nobody

Alibaba's Qwen team released Qwen3-2507, splitting its model line into dedicated instruct and thinking variants. The update brings substantial gains in reasoning, instruction following, and 256K context support, extensible to 1M tokens.

2026-07-29

Labs & Research1 min read

Analysis

Benchmarks, not bargains: China's AI labs reach parity

Chinese AI labs Qwen, DeepSeek, MiniMax, and Kimi have matched or beaten US frontier models in reasoning, coding, and document parsing benchmarks. With aggressive pricing and open-weight strategies, they are no longer just cheap alternatives, they are real competitors.

2026-07-27

AIFeatured4 min read

Real-time interactive worlds, straight from a text prompt

Alibaba just launched an interactive AI world where you can steer the plot

HappyOyster 1.0 generates interactive open worlds from text or images, supporting free movement, plot control, and real-time feedback. Now in gray testing with Android, iOS, and Web SDKs.

2026-07-25

LLMs & Models4 min read

Alibaba Cloud's persistent memory for AI coding

The AI coding partner that finally remembers what you told it last week

Alibaba Cloud's new framework gives AI coding assistants a long-term memory that remembers preferences, history, and project knowledge across sessions. No more starting from zero every time.

2026-07-24

Tools & FrameworksFeatured3 min read

Image generation

Alibaba's latest image model doesn't want to paint. It wants to print your newspaper.

Qwen-Image-3.0 skips the aesthetic polish arms race and targets something harder: generating legible text, dense layouts, and full infographics in one pass. Alibaba just made image generation useful again.

2026-07-23

Benchmarks & Tests2 min read

Agent evaluation

Your AI agent keeps failing? It might be the harness, not the brain

PawBench, an open-source benchmark from the AgentScope team, systematically evaluates models and agent harnesses together. Results show that harness design can swing scores by over 11 points for smaller models, exposing a blind spot in how AI agents are currently judged.

2026-07-22

← PreviousPage 1 / 2 · 22 articlesNext →