1 paper
Yiwen Zhang, Joseph Tung, Ruojin Cai +2
3D foundation models (3DFMs) have recently transformed 3D vision, enabling joint prediction of depths, poses, and point maps directly from images. Yet their ability to reason under…