Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Ant Group released LingBot-Vision: DINO-family vision backbones in 4 sizes, and the 0.3B ViT-L matches DINOv3-7B on NYUv2 depth with ~23x fewer params

Via r/LocalLlama
Monday, Jul 6, 2026 ยท 5:33PM
Summary

Weights, all 4 sizes, Apache-2.0 (ViT-S / ViT-B / ViT-L / ViT-g): https://huggingface.co/collections/robbyant/lingbot-vision Code: https://github.com/robbyant/lingbot-vision Project page: https://technology.robbyant.com/lingbot-vision Self-supervised DINO-family backbone, but the masking is boundary

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories