351 citations · 389 across the 18 of their papers we have counts for
27 papers · 1 filter
Video2Game: Real-time, Interactive, Realistic and Browser-Compatible Environment from a Single Video
Hongchi Xia, Zhi-Hao Lin, Wei-Chiu Ma +1
Creating high-quality and interactive virtual environments, such as games and simulators, often involves complex and costly manual modeling processes. In this paper, we present Vid…
BLINK: Multimodal Large Language Models Can See but Not Perceive
Xingyu Fu, Yushi Hu, Bangzheng Li +7
We introduce Blink, a new benchmark for multimodal language models (LLMs) that focuses on core visual perception abilities not found in other evaluations. Most of the Blink tasks c…
Structure from Duplicates: Neural Inverse Graphics from a Pile of Objects
Tianhang Cheng, Wei-Chiu Ma, Kaiyu Guan +2
Our world is full of identical objects (\emphe.g., cans of coke, cars of same model). These duplicates, when seen together, provide additional and strong cues for us to effectively…
LightSim: Neural Lighting Simulation for Urban Scenes
Ava Pun, Gary Sun, Jingkang Wang +5
Different outdoor illumination conditions drastically alter the appearance of urban scenes, and they can harm the performance of image-based robot perception systems if not seen du…
UltraLiDAR: Learning Compact Representations for LiDAR Completion and Generation
Yuwen Xiong, Wei-Chiu Ma, Jingkang Wang +1
LiDAR provides accurate geometric measurements of the 3D world. Unfortunately, dense LiDARs are very expensive and the point clouds captured by low-beam LiDAR are often sparse. To…
CADSim: Robust and Scalable in-the-wild 3D Reconstruction for Controllable Sensor Simulation
Jingkang Wang, Sivabalan Manivasagam, Yun Chen +5
Realistic simulation is key to enabling safe and scalable development of % self-driving vehicles. A core component is simulating the sensors so that the entire autonomy system can…