multi-agent AI
2 published articles
AI3 min read
Formal mathematics
The bottleneck no one saw in AI math proofs: decomposition, not compute
Nanjing University's ToMap framework achieves state-of-the-art results in full-proof autoformalization by identifying the decomposition step as the critical bottleneck. Using iterative, Pareto-guided evolution of proof decompositions, ToMap lifts joint syntactic-semantic accuracy by 19% on the ProofFlowBench benchmark while reducing test-time costs.
2026-07-30
LLMs & Models4 min read
AI Research
Four minds, one answer: why AI that thinks differently beat the biggest models at humanity's hardest test
PoTRE breaks inference into four agents working in parallel, then reconciles their answers dynamically. It hit 49.92% on Humanity's Last Exam, beating the previous official best. The approach works with fewer tokens than scaled-up baselines.
2026-07-24