activity
20242026
collaborators

11 papers

cs.CV2026

Focus Where It Counts: A Salience-Driven Vision-Language Model for Low Vision Assistance

Jiazhao Liang, Hao Huang, Shuaihang Yuan +8

Vision-language models (VLMs) are rapidly progressing and offer promising capabilities for assistive technologies supporting persons with blindness or low vision. However, existing…

cs.RO2026

Learning Adaptive Safety Margins for Visual Navigation

Junyi Hu, Shuaihang Yuan, Geeta Chandra Raju Bethala +2

Robots in cluttered indoor spaces often fail not because they cannot generate collision-free paths, but because a fixed safety margin is mis-calibrated: conservative margins cause…

cs.RO2026

Immersive Social Interaction with VR and LLM-Assisted Humanoids

Niraj Pudasaini, Geeta Chandra Raju Bethala, Pranav Doma +2

Humanoid robots can extend human presence to remote, constrained, or hazardous environments, but existing teleoperation interfaces often require physically demanding motion trackin…

cs.RO2025

One-shot Adaptation of Humanoid Whole-body Motion with Walking Priors

Hao Huang, Geeta Chandra Raju Bethala, Shuaihang Yuan +4

Whole-body humanoid motion represents a fundamental challenge in robotics, requiring balance, coordination, and adaptability to enable human-like behaviors. However, existing metho…

cs.RO2025

Wavelet Policy: Lifting Scheme for Policy Learning in Long-Horizon Tasks

Hao Huang, Shuaihang Yuan, Geeta Chandra Raju Bethala +3

Policy learning focuses on devising strategies for agents in embodied artificial intelligence systems to perform optimal actions based on their perceived states. One of the key cha…

cs.RO2025

FastMap: Real-Time Semantic Map Completion via Bitwise Masked Modeling

Yijie Deng, Shuaihang Yuan, Congcong Wen +4

Semantic map completion, which predicts the layout of unobserved regions from partial observations, is a critical capability for indoor robot navigation. Existing approaches either…