11 papers
Focus Where It Counts: A Salience-Driven Vision-Language Model for Low Vision Assistance
Jiazhao Liang, Hao Huang, Shuaihang Yuan +8
Vision-language models (VLMs) are rapidly progressing and offer promising capabilities for assistive technologies supporting persons with blindness or low vision. However, existing…
Learning Adaptive Safety Margins for Visual Navigation
Junyi Hu, Shuaihang Yuan, Geeta Chandra Raju Bethala +2
Robots in cluttered indoor spaces often fail not because they cannot generate collision-free paths, but because a fixed safety margin is mis-calibrated: conservative margins cause…
Immersive Social Interaction with VR and LLM-Assisted Humanoids
Niraj Pudasaini, Geeta Chandra Raju Bethala, Pranav Doma +2
Humanoid robots can extend human presence to remote, constrained, or hazardous environments, but existing teleoperation interfaces often require physically demanding motion trackin…
One-shot Adaptation of Humanoid Whole-body Motion with Walking Priors
Hao Huang, Geeta Chandra Raju Bethala, Shuaihang Yuan +4
Whole-body humanoid motion represents a fundamental challenge in robotics, requiring balance, coordination, and adaptability to enable human-like behaviors. However, existing metho…
Wavelet Policy: Lifting Scheme for Policy Learning in Long-Horizon Tasks
Hao Huang, Shuaihang Yuan, Geeta Chandra Raju Bethala +3
Policy learning focuses on devising strategies for agents in embodied artificial intelligence systems to perform optimal actions based on their perceived states. One of the key cha…
FastMap: Real-Time Semantic Map Completion via Bitwise Masked Modeling
Yijie Deng, Shuaihang Yuan, Congcong Wen +4
Semantic map completion, which predicts the layout of unobserved regions from partial observations, is a critical capability for indoor robot navigation. Existing approaches either…