2 citations · 4 across the 4 of their papers we have counts for
4 papers
Single-view 3D Scene Reconstruction with High-fidelity Shape and Texture
Yixin Chen, Junfeng Ni, Nan Jiang +3
Reconstructing detailed 3D scenes from single-view images remains a challenging task due to limitations in existing approaches, which primarily focus on geometric shape recovery, o…
Lightweight In-Context Tuning for Multimodal Unified Models
Yixin Chen, Shuai Zhang, Boran Han +1
In-context learning (ICL) involves reasoning from given contextual examples. As more modalities comes, this procedure is becoming more challenging as the interleaved input modaliti…
3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment
Ziyu Zhu, Xiaojian Ma, Yixin Chen +3
3D vision-language grounding (3D-VL) is an emerging field that aims to connect the 3D physical world with natural language, which is crucial for achieving embodied intelligence. Cu…
Detecting Human-Object Contact in Images
Yixin Chen, Sai Kumar Dwivedi, Michael J. Black +1
Humans constantly contact objects to move and perform tasks. Thus, detecting human-object contact is important for building human-centered artificial intelligence. However, there e…