activity
20242026
collaborators

10 papers

cs.CV2026

Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation

Kam Woh Ng, Jing Yang, Jia Wei Sii +5

Understanding and generating the fine-grained structure of objects -- such as birds with species-specific beaks, wings, and tails -- is a long-standing challenge in computer vision…

cs.CV2026

See Tomorrow, Act Today: Foresight-Driven Autonomous Driving

Bozhou Zhang, Nan Song, Yuang Wang +3

Current end-to-end autonomous driving planners are fundamentally reactive: they condition on historical and present observations to predict future actions. We argue that autonomous…

cs.CV2026

High Dynamic Range 3D Gaussian Splatting via Luminance-Chromaticity Decomposition

Kaixuan Zhang, Minxian Li, Mingwu Ren +2

High Dynamic Range (HDR) 3D reconstruction is pivotal for professional content creation in filmmaking and virtual production. Existing methods typically rely on multi-exposure Low…

cs.CV2026

ImagiDrive: A Unified Imagination-and-Planning Framework for Autonomous Driving

Jingyu Li, Bozhou Zhang, Xin Jin +3

Autonomous driving requires rich contextual comprehension and precise predictive reasoning to navigate dynamic and complex environments safely. Vision-Language Models (VLMs) and Dr…

cs.CV2026

Dynamic Novel View Synthesis in High Dynamic Range

Kaixuan Zhang, Zhipeng Xiong, Minxian Li +3

High Dynamic Range Novel View Synthesis (HDR NVS) seeks to learn an HDR 3D model from Low Dynamic Range (LDR) training images captured under conventional imaging conditions. Curren…

cs.CV2025

ShapeCraft: LLM Agents for Structured, Textured and Interactive 3D Modeling

Shuyuan Zhang, Chenhan Jiang, Zuoou Li +1

3D generation from natural language offers significant potential to reduce expert manual modeling efforts and enhance accessibility to 3D assets. However, existing methods often yi…