蚂蚁集团开源机器人视觉基础模型LingBot-Vision
I'll always root for a team that open-sources its best work, and Robbyant just d…
做机器人或具身智能的同学注意了,蚂蚁开源了视觉基础模型,数据效率极高且无需标注,直接拿来就能用,赶紧试试深度估计效果。
I'll always root for a team that open-sources its best work, and Robbyant just did it properly.
Robbyant, Ant Group's embodied-AI company, released LingBot-Vision, a vision foundation model for robots, and the part I love is the data. They trained it on 161M images, filtered down from 2B raw ones and mostly pulled straight from the open web, with no human labels, no edge detectors, no depth sensors anywhere in the loop. It learns the exact edges of objects from raw pixels. That's roughly a tenth of the data DINOv3 saw, and under a third of the training.
And it shows in the results. On depth, working out how far away things are, the 1B model edges out a 7B on NYU-Depth. It also powers LingBot-Depth 2.0, which reads the surfaces cameras usually choke on, glass and mirrors, and halves indoor depth error.
LingBot-Vision is fully open. Weights from the 1.1B flagship down to a tiny 21M version, code, and the paper. This is the timeline I want more of. @robbyant_brain
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力