SevenTnewS

AI coding agents

12 published articles

AI IDEs4 min read

AI coding security

Qoder moves code security from the release gate to the keystroke

Qoder Desktop and CLI now run security checks at three depths, from live keystroke screening to cross-file data-flow analysis before a merge. Alibaba Cloud reports roughly 60% better vulnerability detection and a 35% to 45% drop in security feedback during code review.

2026-08-13

AI Agents5 min read

AI Agents & DevOps: STAROps inside Qoder

A one-sentence question in Qoder ends the 40-minute root-cause hunt

Alibaba Cloud's STAROps plugin brings natural-language root-cause diagnosis into the Qoder IDE. Its demo takes a P95 spike from under 60ms to 1.9 seconds, returns an evidence chain at 80% confidence, and ends with an auto-created merge request. Traditional troubleshooting ran 40-plus minutes across five platforms.

2026-08-09

AI IDEs6 min read

Agentic IDEs: betting on task boundaries

Alibaba says Qoder 1.0 cut agent input tokens 40%. Alibaba ran the test

Qoder 1.0 rearchitects the AI coding IDE around task boundaries, claiming a 40% cut in agent input tokens and a 22% drop in dissatisfaction on Alibaba's own three-day A/B test. No independent replication has been published yet.

2026-08-09

Meta AIFeatured4 min read

Meta AI

Muse Code: Meta's crash-resilient bet in the coding agent race

Meta's Muse Code brings a crash-surviving terminal coding agent to macOS and Linux, powered by Muse Spark 1.2, a model co-trained with its own harness for long-horizon tasks. Meta says larger, more capable models are already on the way.

2026-08-06

AI Agents3 min read

AI Development

Coding agents are leaving your local machine: Alibaba and Mistral go remote

Alibaba Cloud's Qoder introduces remote delegation for coding agents, letting developers offload long-running tasks to cloud sandboxes. The feature, paired with a new knowledge engine, pits Qoder directly against Mistral's remote agents in a race to decouple AI-assisted development from the local terminal.

2026-08-03

AI AgentsFeatured2 min read

AI Agents

Nobody Can Say What Their AI Coding Agents Actually Did Today

Enterprises spending heavily on AI coding agents can't see what those agents actually did, the benchmarks measuring their improvement may be overfit to public test sets, and the code they write isn't reviewed by default. Three separate problems with the same root cause: adoption outran the tooling to observe it.

2026-07-30

AI AgentsFeatured4 min read

Agent Observability

Your AI coding agents are burning millions, and nobody knows where

Enterprises are spending heavily on AI coding agents with zero visibility into their actual behavior. Alibaba Cloud's open-source LoongSuite Pilot aims to fill that gap, but the core issue runs deeper: nobody can track what these agents are doing.

2026-07-26

Apple Intelligence4 min read

WWDC 2025

Apple's walled garden just got a gate with an AI-shaped key

Xcode 27 brings agentic coding with Anthropic, Google, and OpenAI models, plus new intelligence frameworks that let apps use models from Apple, Claude, Gemini, and others. Core AI lets developers run full-scale LLMs on-device.

2026-07-25

AI AgentsFeatured2 min read

AI Coding Agents

How an Anthropic ban backfired and drove OpenCode to $40M ARR

OpenCode, an open-source AI coding agent, reached 4.6 million weekly active users and $40M ARR after an Anthropic clampdown on Claude usage pushed developers toward a model-agnostic alternative.

2026-07-24

Qwen / AlibabaFeatured4 min read

Closed-loop coding agents

Alibaba built a coding model that learns from its own users. The numbers are hard to ignore.

Alibaba's new coding model Qwen-Coder-Qoder beats Cursor Composer-1 on the Qoder Bench benchmark. Production metrics show a 3.85% code retention increase, a 61.5% drop in tool errors, and a 14.5% reduction in token use. The model trains on real agent traces through a rewarder-attacker framework against reward hacking.

2026-07-23

AIFeatured4 min read

Coding agents

Kimi K3 edges out claude fable 5 and gpt 5.6 sol on next.js code gen benchmark

Kimi K3 ties for first at 92% success rate on Next.js code tasks, finishing in under 200 seconds. AGENTS.md documentation erases gaps between top models and mid-tier ones, pushing all of them to 96%.

2026-07-18

AI IDEsFeatured3 min read

Mission control, not remote terminal

Cursor's iOS beta turns your phone into a coding agent dispatcher

Cursor's iOS app in public beta lets developers launch AI coding agents from their phone, bridge local and cloud workflows, and get push notifications when PRs are ready. It rethinks mobile's role in software development beyond incident response.

2026-07-14