2 papers
cs.CV2026
SAM 3D Animal: Promptable Animal 3D Reconstruction from Images in the Wild
Xuyi Hu, Jin Lyu, Jiuming Liu +4
3D animal reconstruction in the wild remains challenging due to large species variation, frequent occlusions, and the prevalence of multi-animal scenes, while existing methods pred…
cs.CV2026
ForgeVLA: Federated Vision-Language-Action Learning without Language Annotations
Yuhao Zhou, Yunpeng Zhu, Yang Zhou +7
Vision-Language-Action (VLA) models hold great promise for general-purpose robotic intelligence, yet scaling up such models is severely bottlenecked by the high cost of acquiring a…