visual reasoning
3 published articles
Research
Beacon: your tool-using AI model is making easy questions harder
A new paper from KlingTeam measures when multimodal models actually need tools and when tools hurt. The proposed Beacon model, trained with necessity-aware rewards, improves both accuracy and tool discipline.
2026-08-09
AI Labs & Research
An AI that can see its own mistakes and undo them
A training-free AI sampler that generates images and text together, retracting decisions when cross-modal signals contradict each other. The paper's CO₂Jump outperforms existing methods and suggests a structural fix for a blind spot in multimodal AI.
2026-07-17
Generative AI
An AI that can see its own mistakes and undo them mid-generation
Google introduces CO2Jump, a training-free sampler for joint text and image generation. It uses a self-correcting Markov jump process where each modality's confidence scores guide the other's updates in real time, catching cross-modal errors mid-generation.
2026-07-14