Qwen
22 published articles
AI Music Generation: Qwen's answer to prompt drift
Music 3.0 swaps the one-tag prompt for a timeline that keeps AI songs on track
AI tracks tend to drift from the prompt as they unfold: instruments drop out, emotion flattens, the vocal style comes and goes. Music 3.0 swaps the one-tag description for a time-sequential Structured Caption, backed by an 8B/0.6B Hybrid-LM that splits structure from detail.
2026-08-21
Multimodal / Open Source
Qwen plugin pack gives your coding agent eyes, hands, and video memory
Alibaba's Qwen team released Qwen-MM-Plugins, an Apache-2.0 toolkit that gives coding agents native multimodal skills: dynamic-resolution image and video reading, OCR, grounding, ASR, long-video memory, video generation, and thin-client control of Blender and FreeCAD. One script installs it across Claude Code, Codex, Qoder, OpenClaw, Qwen Code, and Gemini CLI.
2026-08-15
AI Economics
Alibaba's $18 plan runs Qwen and DeepSeek inside Claude Code
Alibaba Cloud's Token Plan Individual bundles Qwen and third-party models such as DeepSeek into a single credit pool, from $6 a month, usable inside Claude Code and Cursor. The company claims roughly 40% savings over pay-as-you-go, but the plan only works from Singapore and only inside approved tools.
2026-08-14
Cloud & LLM Security
Alibaba Cloud's AI gateway blocks prompt attacks with a 200, not a 403
A hands-on walkthrough of Alibaba Cloud's AI Gateway shows how authentication, guardrails and PII masking work as one pipeline. The gotchas: blocks return HTTP 200, key enforcement needs a manual toggle, and restored data can end up in logs.
2026-08-09
AI MUSIC
The melody-first trick that lets Qwen-Music beat Suno on 13 metrics
Qwen-Music separates melody planning from full-song generation via Melody-CoT. The 3-billion-parameter model, trained on 5 million+ hours of multilingual data, achieves state-of-the-art results against Suno V5 and MiniMax Music 2.6 in both objective metrics and human preference.
2026-07-31
GUI Agents
Alibaba's Qwen-UI-Agent scores higher on real phones than in sandboxes
Alibaba Tongyi Lab's Qwen-UI-Agent claims state-of-the-art mobile scores (97.5% AndroidDaily) and competitive computer-use results against Opus 4.8, Gemini 3.1 Pro, and GPT-5.6 Sol. Its real-device benchmark score beats its sandbox score, while 40% partial progress on OSWorld-v2 is the honest limit.
2026-07-31
Qwen / Alibaba
Alibaba splits its flagship AI model in two because one size fits nobody
Alibaba's Qwen team released Qwen3-2507, splitting its model line into dedicated instruct and thinking variants. The update brings substantial gains in reasoning, instruction following, and 256K context support, extensible to 1M tokens.
2026-07-29
Analysis
Benchmarks, not bargains: China's AI labs reach parity
Chinese AI labs Qwen, DeepSeek, MiniMax, and Kimi have matched or beaten US frontier models in reasoning, coding, and document parsing benchmarks. With aggressive pricing and open-weight strategies, they are no longer just cheap alternatives, they are real competitors.
2026-07-27
Real-time interactive worlds, straight from a text prompt
Alibaba just launched an interactive AI world where you can steer the plot
HappyOyster 1.0 generates interactive open worlds from text or images, supporting free movement, plot control, and real-time feedback. Now in gray testing with Android, iOS, and Web SDKs.
2026-07-25
Alibaba Cloud's persistent memory for AI coding
The AI coding partner that finally remembers what you told it last week
Alibaba Cloud's new framework gives AI coding assistants a long-term memory that remembers preferences, history, and project knowledge across sessions. No more starting from zero every time.
2026-07-24
Image generation
Alibaba's latest image model doesn't want to paint. It wants to print your newspaper.
Qwen-Image-3.0 skips the aesthetic polish arms race and targets something harder: generating legible text, dense layouts, and full infographics in one pass. Alibaba just made image generation useful again.
2026-07-23
Agent evaluation
Your AI agent keeps failing? It might be the harness, not the brain
PawBench, an open-source benchmark from the AgentScope team, systematically evaluates models and agent harnesses together. Results show that harness design can swing scores by over 11 points for smaller models, exposing a blind spot in how AI agents are currently judged.
2026-07-22