10 papers
Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation
Kam Woh Ng, Jing Yang, Jia Wei Sii +5
Understanding and generating the fine-grained structure of objects -- such as birds with species-specific beaks, wings, and tails -- is a long-standing challenge in computer vision…
See Tomorrow, Act Today: Foresight-Driven Autonomous Driving
Bozhou Zhang, Nan Song, Yuang Wang +3
Current end-to-end autonomous driving planners are fundamentally reactive: they condition on historical and present observations to predict future actions. We argue that autonomous…
High Dynamic Range 3D Gaussian Splatting via Luminance-Chromaticity Decomposition
Kaixuan Zhang, Minxian Li, Mingwu Ren +2
High Dynamic Range (HDR) 3D reconstruction is pivotal for professional content creation in filmmaking and virtual production. Existing methods typically rely on multi-exposure Low…
ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving
Jingyu Li, Bozhou Zhang, Xin Jin +3
Autonomous driving requires rich contextual comprehension and precise predictive reasoning to navigate dynamic and complex environments safely. Vision-Language Models (VLMs) and Dr…
Dynamic Novel View Synthesis in High Dynamic Range
Kaixuan Zhang, Zhipeng Xiong, Minxian Li +3
High Dynamic Range Novel View Synthesis (HDR NVS) seeks to learn an HDR 3D model from Low Dynamic Range (LDR) training images captured under conventional imaging conditions. Curren…
ShapeCraft: LLM Agents for Structured, Textured and Interactive 3D Modeling
Shuyuan Zhang, Chenhan Jiang, Zuoou Li +1
3D generation from natural language offers significant potential to reduce expert manual modeling efforts and enhance accessibility to 3D assets. However, existing methods often yi…