collaborators

6 papers

cs.SD2026

NeuralMUSIC: A Hybrid Neural-Subspace Framework for Robot Sound Source Localization

Yizhuo Yang, Junqiao Fan, Shenghai Yuan +1

Reliable sound source localization is fundamental to robot audition, enabling autonomous robots to perceive spatial cues and operate effectively in dynamic environments. Classical…

cs.RO2026

GIVE: Grounding Human Gestures in Vision-Language-Action Models

Pengfei Liu, Gen Li, Junqiao Fan +4

Human communication is inherently multimodal, where language is often accompanied by non-verbal cues such as gestures to convey intentions. However, current Vision-Language-Action…

cs.CV2026

M4Human: A Large-Scale Multimodal mmWave Radar Benchmark for Human Mesh Reconstruction

Junqiao Fan, Yunjiao Zhou, Yizhuo Yang +6

Human mesh reconstruction (HMR) provides direct insights into body-environment interaction, which enables various immersive applications. While existing large-scale HMR datasets re…

cs.CV2025

SMamDiff: Spatial Mamba for Stochastic Human Motion Prediction

Junqiao Fan, Pengfei Liu, Haocong Rao

With intelligent room-side sensing and service robots widely deployed, human motion prediction (HMP) is essential for safe, proactive assistance. However, many existing HMP methods…

cs.CV2025

mmPred: Radar-based Human Motion Prediction in the Dark

Junqiao Fan, Haocong Rao, Jiarui Zhang +2

Existing Human Motion Prediction (HMP) methods based on RGB-D cameras are sensitive to lighting conditions and raise privacy concerns, limiting their real-world applications such a…

cs.CV2025

Generative Dataset Distillation using Min-Max Diffusion Model

Junqiao Fan, Yunjiao Zhou, Min Chang Jordan Ren +1

In this paper, we address the problem of generative dataset distillation that utilizes generative models to synthesize images. The generator may produce any number of images under…