Ant Group Open-Sources LingBot-Vision: 1B Boundary-Centric Spatial Perception Model

Loading…

Ant Group's robotics division RobbyAnt has released LingBot-Vision, a 1-billion parameter vision foundation model specifically optimized for dense spatial perception with a focus on boundary detection and object edge understanding. Unlike general-purpose vision encoders, LingBot-Vision is designed for downstream robotics and manipulation tasks where precise spatial boundaries — not just object classification — determine whether an action succeeds or fails. At 1B parameters, it is sized for deployment on edge hardware and embedded robot controllers rather than cloud inference, which is a deliberate design choice for real-world robotics applications. The open-source release makes it directly usable by robotics developers who need a compact, boundary-aware vision backbone without training from scratch. This is one of the more practically targeted open vision releases for the robotics and embodied AI community in recent months.