14 papers
Learning Adaptive Safety Margins for Visual Navigation
Junyi Hu, Shuaihang Yuan, Geeta Chandra Raju Bethala +2
Robots in cluttered indoor spaces often fail not because they cannot generate collision-free paths, but because a fixed safety margin is mis-calibrated: conservative margins cause…
VTaMo: Video-Text Alignment Model for Sign Language Translation
Junyi Hu, Zhewen He, Haomian Huang +2
Sign language translation (SLT) converts continuous sign videos into spoken language text. Gloss-free approaches leverage pre-trained visual encoders and language models but rely o…
Immersive Social Interaction with VR and LLM-Assisted Humanoids
Niraj Pudasaini, Geeta Chandra Raju Bethala, Pranav Doma +2
Humanoid robots can extend human presence to remote, constrained, or hazardous environments, but existing teleoperation interfaces often require physically demanding motion trackin…
VCS-SLAM: Geometry-Validated Semantic Evidence Fusion for 3D Gaussian SLAM
Raman Jha, Shuaihang Yuan, Yi Fang
Visual SLAM performance often deteriorates in complex real-world applications. Semantic 3D Gaussian SLAM commonly fuses 2D semantic priors into a persistent 3D map using uniform op…
SignNet-1M: Large-Scale Multilingual Sign Language Video Dataset with Downstream Benchmarks
Zhewen He, Junyi Hu, Haomian Huang +3
Sign language models are typically trained on datasets captured under constrained conditions, with limited viewpoint, background, and signer-identity diversity, leading to poor rob…
One-shot Adaptation of Humanoid Whole-body Motion with Walking Priors
Hao Huang, Geeta Chandra Raju Bethala, Shuaihang Yuan +4
Whole-body humanoid motion represents a fundamental challenge in robotics, requiring balance, coordination, and adaptability to enable human-like behaviors. However, existing metho…