AI Agents5 min read
Autonomous coding agents · arXiv preprint
Harness-of-Harness claims a 52.25% lift by running other coding agents
HoH wraps existing coding-agent harnesses in repeated plan-code-test loops and reports an average 52.25% relative gain across three benchmark suites. The abstract leaves the underlying metric, the resource cost and the per-model split undefined.
2026-09-25