2 papers
cs.CV2026
Fast SceneScript: Fast and Accurate Language-Based 3D Scene Understanding via Multi-Token Prediction
Ruihong Yin, Xuepeng Shi, Oleksandr Bailo +2
Recent perception-generalist approaches based on language models have achieved state-of-the-art results across diverse tasks, including 3D scene layout estimation and 3D object det…
cs.CV2025
SDFit: 3D Object Pose and Shape by Fitting a Morphable SDF to a Single Image
Dimitrije AntiÄ, Georgios Paschalidis, Shashank Tripathi +3
Recovering 3D object pose and shape from a single image is a challenging and ill-posed problem. This is due to strong (self-)occlusions, depth ambiguities, the vast intra- and inte…