3 papers
cs.CV2026
GuideGround: VLM-guided Semantic Understanding and Viewpoint-aware Reasoning for 3D Visual Grounding
Yiwen Wang, Yuyang Deng, Yihao Long +1
3D visual grounding aims to localize the target object in a 3D scene from a natural language query, requiring both fine-grained semantic understanding and viewpoint-dependent spati…
cs.CV2026
GTSR: Subsurface Scattering Awared 3D Gaussians for Translucent Surface Reconstruction
Youwen Yuan, Xi Zhao
Reconstructing translucent objects from multi-view images is a difficult problem. Previously, researchers have used differentiable path tracing and the neural implicit field, which…
cs.CV2025
PBR3DGen: A VLM-guided Mesh Generation with High-quality PBR Texture
Xiaokang Wei, Bowen Zhang, Xianghui Yang +4
Generating high-quality physically based rendering (PBR) materials is important to achieve realistic rendering in the downstream tasks, yet it remains challenging due to the intert…