collaborators

5 papers

cs.CV2026

CapNav: Benchmarking Vision Language Models on Capability-conditioned Indoor Navigation

Xia Su, Ruiqi Chen, Benlin Liu +4

Vision-Language Models (VLMs) have shown remarkable progress in Vision-Language Navigation (VLN), offering new possibilities for navigation decision-making that could benefit both…

cs.HC2025

DepthScape: Authoring 2.5D Designs via Depth Estimation, Semantic Understanding, and Geometry Extraction

Xia Su, Cuong Nguyen, Matheus A. Gadelha +1

2.5D effects, such as occlusion and perspective foreshortening, enhance visual dynamics and realism by incorporating 3D depth cues into 2D designs. However, creating such effects r…

cs.HC2025

SonoCraftAR: Towards Supporting Personalized Authoring of Sound-Reactive AR Interfaces by Deaf and Hard of Hearing Users

Jaewook Lee, Davin Win Kyi, Leejun Kim +5

Augmented reality (AR) has shown promise for supporting Deaf and hard-of-hearing (DHH) individuals by captioning speech and visualizing environmental sounds, yet existing systems d…

cs.HC2025

ImaginateAR: AI-Assisted In-Situ Authoring in Augmented Reality

Jaewook Lee, Filippo Aleotti, Diego Mazala +9

While augmented reality (AR) enables new ways to play, tell stories, and explore ideas rooted in the physical world, authoring personalized AR content remains difficult for non-exp…

cs.HC2025

Accessibility Scout: Personalized Accessibility Scans of Built Environments

William Huang, Xia Su, Jon E. Froehlich +1

Assessing the accessibility of unfamiliar built environments is critical for people with disabilities. However, manual assessments, performed by users or their personal health prof…