Labs & Research
OpenAI, Anthropic, Google DeepMind, Meta AI and publications.
114 published articles
Qwen / Alibaba Cloud
Qwen2.5-Omni outperforms Gemini-1.5-Pro on OmniBench, fits under 12GB
Alibaba's open-source Qwen2.5-Omni outscored Gemini-1.5-Pro on OmniBench and topped the MMAU audio reasoning leaderboard. Quantized builds cut VRAM below 12GB and MNN support brings real-time voice chat to phones.
2026-08-07
Funding Round
Mistral AI's 1.7 billion euro bet on the factory floor
Mistral AI's Series C, led by ASML, raises 1.7 billion euros at an 11.7 billion valuation. The partnership ties the AI startup to semiconductor manufacturing, marking a strategic pivot from general-purpose models to industrial applications.
2026-08-06
AI Policy
Why Anthropic handed its AI policy to a former Supreme Court justice
Former California Supreme Court Justice Tino Cuéllar joins Anthropic as its first Chief Global Affairs Officer. His hiring signals that the AI governance fight will be decided in capitals, and that Anthropic is staffing up to be in those rooms.
2026-08-06
Enterprise AI
Forget the model race: Alibaba is automating security operations
Qwen3.8-Max grabbed the headlines, but Alibaba's real move is quieter: AI agents that run security operations inside its cloud console. We break down the SecOps Agent, the Qwen-powered fraud forensics, and the lock-in strategy behind the scores.
2026-08-06
Qwen3.8-Max open weights land next week
Alibaba's most powerful model ever is going open source
Qwen3.8-Max, Alibaba's first open-weight Max-class model at 2.4 trillion parameters, hits Hugging Face and ModelScope next week. The launch reframes the open-source question: what happens when the frontier's biggest weights are free to download and test?
2026-08-06
Meta AI
Muse Code: Meta's crash-resilient bet in the coding agent race
Meta's Muse Code brings a crash-surviving terminal coding agent to macOS and Linux, powered by Muse Spark 1.2, a model co-trained with its own harness for long-horizon tasks. Meta says larger, more capable models are already on the way.
2026-08-06
AI consolidation
Reve's researchers are headed to OpenAI; its products stay put
OpenAI's investment in the independent lab Reve brings a move of research talent: part of Reve's AI Research team joins OpenAI's multimodal push while Reve's products stay in place. No amount was disclosed.
2026-08-05
AI Safety: 3B Classifier, Apache 2.0 Weights
Mistral's Shieldstral puts your moderation policy in the prompt, not the weights
Shieldstral frames moderation as a binary question: an instruction, a yes/no query, and the content to judge. Mistral says the 3B model matches open guardrails up to seven times its size on text safety, with Apache 2.0 weights that run on one 16GB GPU.
2026-08-05
Local-language AI meets the Japanese enterprise
Sakana Namazu bets on keigo against the strongest AI era
Sakana AI introduced Namazu, a model specialized for Japanese business documents, email and keigo. It arrives as Asian enterprises report that scarce, low-quality local-language models are holding back AI adoption.
2026-08-03
Open-source AI
Qwen3.8-Max beat 458 human teams in 24 hours, working alone
Qwen3.8-Max beat 458 of 526 human teams in a 24-hour contest while working alone, and Alibaba will open-source its weights next week. Every number is self-reported so far, which is exactly why the autonomy claims deserve scrutiny.
2026-08-03
Product strategy: Sakana's four-product sprint
Sakana AI bets against the 'strongest AI' era: orchestration, shipped fast
Sakana AI released Sakana Chat, Marlin, Fugu and Translate in quick succession. Head of product Sota Omura explains why shipping fast is the AI-era playbook, and why the company is betting against a single 'strongest AI'.
2026-08-03
Frontier model access
Anthropic's smartest Claude is the one you can't use
Anthropic launched Claude Fable 5 for everyone and kept Claude Mythos 5 for vetted partners. Same model class, two access doors. The benchmarks matter less than the routing, and the routing is how frontier AI ships from here.
2026-08-03