6 papers
NeuralMUSIC: A Hybrid Neural-Subspace Framework for Robot Sound Source Localization
Yizhuo Yang, Junqiao Fan, Shenghai Yuan +1
Reliable sound source localization is fundamental to robot audition, enabling autonomous robots to perceive spatial cues and operate effectively in dynamic environments. Classical…
GIVE: Grounding Human Gestures in Vision-Language-Action Models
Pengfei Liu, Gen Li, Junqiao Fan +4
Human communication is inherently multimodal, where language is often accompanied by non-verbal cues such as gestures to convey intentions. However, current Vision-Language-Action…
M4Human: A Large-Scale Multimodal mmWave Radar Benchmark for Human Mesh Reconstruction
Junqiao Fan, Yunjiao Zhou, Yizhuo Yang +6
Human mesh reconstruction (HMR) provides direct insights into body-environment interaction, which enables various immersive applications. While existing large-scale HMR datasets re…
SMamDiff: Spatial Mamba for Stochastic Human Motion Prediction
Junqiao Fan, Pengfei Liu, Haocong Rao
With intelligent room-side sensing and service robots widely deployed, human motion prediction (HMP) is essential for safe, proactive assistance. However, many existing HMP methods…
mmPred: Radar-based Human Motion Prediction in the Dark
Junqiao Fan, Haocong Rao, Jiarui Zhang +2
Existing Human Motion Prediction (HMP) methods based on RGB-D cameras are sensitive to lighting conditions and raise privacy concerns, limiting their real-world applications such a…
Generative Dataset Distillation using Min-Max Diffusion Model
Junqiao Fan, Yunjiao Zhou, Min Chang Jordan Ren +1
In this paper, we address the problem of generative dataset distillation that utilizes generative models to synthesize images. The generator may produce any number of images under…