1M context
3 published articles
AI Models | NVIDIA Nemotron 3.5 Lightning
Why Nemotron 3.5 Lightning bets most agent steps don't need a big model
Nvidia's new open model runs locally with 3B active parameters and a 1M-token context, built for agents that stay running. Nvidia claims up to 4x throughput and 30% faster task completion than comparable open models.
2026-08-14
Model Release
Anthropic Unveils Claude Sonnet 5: A Hybrid Reasoning Model for Real-Time Agents and High-Volume Work
Anthropic's Claude Sonnet 5 delivers hybrid reasoning with a 1M context window, excelling at coding, agentic workflows, and enterprise tasks. Available at an introductory price of $2 per million input tokens, promising multi-day coding projects compressed into hours.
2026-07-09
AI Labs & Research
MiniMax's M3 just wrote its own CUDA kernel, and opened the code
MiniMax M3 scores 83.5 on BrowseComp, edges past Opus 4.7, and handles up to 1M tokens natively. In a remarkable autonomy test, it self-optimized a GPU kernel from 7.6% to 71.3% peak utilization without human intervention.
2026-07-09