News
AI Summary
7 Jul 202622 Muharram 1448 AH
Ant Group's LingBot-Vision Claims 12 World Firsts: 1.1B Parameter Model Beats 7B DINOv3

Ant Group's LingBot-Vision Claims 12 World Firsts: 1.1B Parameter Model Beats 7B DINOv3

Ant Group has launched LingBot-Depth 2.0, achieving 12 world-first results across benchmark datasets. This model utilizes the LingBot-Vision foundation model, the first spatial-native vision model designed for embodied AI, released as open source. LingBot-Vision is trained for first-person spatial perception, embedding geometric understanding from the start. With approximately 1.1 billion parameters, it significantly outperforms DINOv3, which has 7 billion parameters, especially in depth estimation tasks.

Follow these topics

Sign in to follow the topics that matter to you

Sign in to follow

This summary is generated with AI and receives periodic editorial review. Refer to the original source for full details.

0
0 reading now

Insight Score

Rate to unlock

Sign in to react, rate, and save. Sign In