collaborators

17 papers

cs.CV2026

Decoupling Cross-Modality Manifold Discrepancy: Leveraging Visible Diffusion Priors for Infrared Super-Resolution

Yunpeng Hua, Hongwei Yu, Jiawei Li +3

Infrared image super-resolution (IISR) mitigates the limitations imposed by low spatial resolution. Existing methods have recognized that IISR should preserve consistency in global…

cs.RO2026

PACT: Proactive Asking for Continual Task Assistance in Human-Robot Collaboration

Chengbo He, Sheng Li, Chenyang Ma +6

Robotic assistants in long-term human-robot collaboration need to assist users under partial observations while leveraging cross-day interaction history. However, human traits and…

cs.CV2026

Unposed-to-3D: Learning Simulation-Ready Vehicles from Real-World Images

Hongyuan Liu, Bochao Zou, Qiankun Liu +10

Creating realistic and simulation-ready 3D assets is crucial for autonomous driving research and virtual environment construction. However, existing 3D vehicle generation methods a…

cs.CV2026

Video-Only ToM: Enhancing Theory of Mind in Multimodal Large Language Models

Siqi Liu, Xinyang Li, Bochao Zou +3

As large language models (LLMs) continue to advance, there is increasing interest in their ability to infer human mental states and demonstrate a human-like Theory of Mind (ToM). M…

cs.CV2025

XYZCylinder: Towards Compatible Feed-Forward 3D Gaussian Splatting for Driving Scenes via Unified Cylinder Lifting Method

Haochen Yu, Qiankun Liu, Hongyuan Liu +4

Feed-forward paradigms for 3D reconstruction have become a focus of recent research, which learn implicit, fixed view transformations to generate a single scene representation. How…

cs.CV2025

MVSMamba: Multi-View Stereo with State Space Model

Jianfei Jiang, Qiankun Liu, Hongyuan Liu +4

Robust feature representations are essential for learning-based Multi-View Stereo (MVS), which relies on accurate feature matching. Recent MVS methods leverage Transformers to capt…