AI coding agents
12 published articles
AI coding security
Qoder moves code security from the release gate to the keystroke
Qoder Desktop and CLI now run security checks at three depths, from live keystroke screening to cross-file data-flow analysis before a merge. Alibaba Cloud reports roughly 60% better vulnerability detection and a 35% to 45% drop in security feedback during code review.
2026-08-13
AI Agents & DevOps: STAROps inside Qoder
A one-sentence question in Qoder ends the 40-minute root-cause hunt
Alibaba Cloud's STAROps plugin brings natural-language root-cause diagnosis into the Qoder IDE. Its demo takes a P95 spike from under 60ms to 1.9 seconds, returns an evidence chain at 80% confidence, and ends with an auto-created merge request. Traditional troubleshooting ran 40-plus minutes across five platforms.
2026-08-09
Agentic IDEs: betting on task boundaries
Alibaba says Qoder 1.0 cut agent input tokens 40%. Alibaba ran the test
Qoder 1.0 rearchitects the AI coding IDE around task boundaries, claiming a 40% cut in agent input tokens and a 22% drop in dissatisfaction on Alibaba's own three-day A/B test. No independent replication has been published yet.
2026-08-09
Meta AI
Muse Code: Meta's crash-resilient bet in the coding agent race
Meta's Muse Code brings a crash-surviving terminal coding agent to macOS and Linux, powered by Muse Spark 1.2, a model co-trained with its own harness for long-horizon tasks. Meta says larger, more capable models are already on the way.
2026-08-06
AI Development
Coding agents are leaving your local machine: Alibaba and Mistral go remote
Alibaba Cloud's Qoder introduces remote delegation for coding agents, letting developers offload long-running tasks to cloud sandboxes. The feature, paired with a new knowledge engine, pits Qoder directly against Mistral's remote agents in a race to decouple AI-assisted development from the local terminal.
2026-08-03
AI Agents
Nobody Can Say What Their AI Coding Agents Actually Did Today
Enterprises spending heavily on AI coding agents can't see what those agents actually did, the benchmarks measuring their improvement may be overfit to public test sets, and the code they write isn't reviewed by default. Three separate problems with the same root cause: adoption outran the tooling to observe it.
2026-07-30
Agent Observability
Your AI coding agents are burning millions, and nobody knows where
Enterprises are spending heavily on AI coding agents with zero visibility into their actual behavior. Alibaba Cloud's open-source LoongSuite Pilot aims to fill that gap, but the core issue runs deeper: nobody can track what these agents are doing.
2026-07-26
WWDC 2025
Apple's walled garden just got a gate with an AI-shaped key
Xcode 27 brings agentic coding with Anthropic, Google, and OpenAI models, plus new intelligence frameworks that let apps use models from Apple, Claude, Gemini, and others. Core AI lets developers run full-scale LLMs on-device.
2026-07-25
AI Coding Agents
How an Anthropic ban backfired and drove OpenCode to $40M ARR
OpenCode, an open-source AI coding agent, reached 4.6 million weekly active users and $40M ARR after an Anthropic clampdown on Claude usage pushed developers toward a model-agnostic alternative.
2026-07-24
Closed-loop coding agents
Alibaba built a coding model that learns from its own users. The numbers are hard to ignore.
Alibaba's new coding model Qwen-Coder-Qoder beats Cursor Composer-1 on the Qoder Bench benchmark. Production metrics show a 3.85% code retention increase, a 61.5% drop in tool errors, and a 14.5% reduction in token use. The model trains on real agent traces through a rewarder-attacker framework against reward hacking.
2026-07-23
Coding agents
Kimi K3 edges out claude fable 5 and gpt 5.6 sol on next.js code gen benchmark
Kimi K3 ties for first at 92% success rate on Next.js code tasks, finishing in under 200 seconds. AGENTS.md documentation erases gaps between top models and mid-tier ones, pushing all of them to 96%.
2026-07-18
Mission control, not remote terminal
Cursor's iOS beta turns your phone into a coding agent dispatcher
Cursor's iOS app in public beta lets developers launch AI coding agents from their phone, bridge local and cloud workflows, and get push notifications when PRs are ready. It rethinks mobile's role in software development beyond incident response.
2026-07-14