Vision & Diffusion4 min read
Computer Vision
CABiNet stays within 2 points of YOLO26x at an eighth of the compute
The VDD Semantic Segmentation Model Zoo brings YOLO26 and CABiNet models trained on varied drone footage to Hugging Face. YOLO26x-sem leads at 78.83% mIoU, but CABiNet-Large's 77.76% at 54.8 GFLOPs makes the efficiency case.
2026-08-07