collaborators

8 papers

cs.RO2025

Unveiling the Impact of Data and Model Scaling on High-Level Control for Humanoid Robots

Yuxi Wei, Zirui Wang, Kangning Yin +3

Data scaling has long remained a critical bottleneck in robot learning. For humanoid robots, human videos and motion data are abundant and widely available, offering a free and lar…

cs.CV2025

Communication-Efficient Multi-Agent 3D Detection via Hybrid Collaboration

Yue Hu, Juntong Peng, Yunqiao Yang +1

Collaborative 3D detection can substantially boost detection performance by allowing agents to exchange complementary information. It inherently results in a fundamental trade-off…

cs.CV2025

InfiniCube: Unbounded and Controllable Dynamic 3D Driving Scene Generation with World-Guided Video Models

Yifan Lu, Xuanchi Ren, Jiawei Yang +8

We present InfiniCube, a scalable method for generating unbounded dynamic 3D driving scenes with high fidelity and controllability. Previous methods for scene generation either suf…

cs.RO2025

BeliefMapNav: 3D Voxel-Based Belief Map for Zero-Shot Object Navigation

Zibo Zhou, Yue Hu, Lingkai Zhang +2

Zero-shot object navigation (ZSON) allows robots to find target objects in unfamiliar environments using natural language instructions, without relying on pre-built maps or task-sp…

cs.CV2025

Towards Collaborative Autonomous Driving: Simulation Platform and End-to-End System

Genjia Liu, Yue Hu, Chenxin Xu +8

Vehicle-to-everything-aided autonomous driving (V2X-AD) has a huge potential to provide a safer driving solution. Despite extensive researches in transportation and communication t…

cs.CV2025

Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models

Yuzhu Cai, Sheng Yin, Yuxi Wei +5

The burgeoning landscape of text-to-image models, exemplified by innovations such as Midjourney and DALLE 3, has revolutionized content creation across diverse sectors. However, th…