3 papers
cs.RO2026
Beyond Object Selection:Markerless Gaze-based Robot Placement at Arbitrary Position
Yuzhi Lai, William Marx, Shenghai Yuan +3
Gaze-based assistive manipulation typically supports object selection, while arbitrary-position placement requires accurate spatial alignment between the headset and robot. However…
cs.CV2026
Can Multimodal Large Language Models Understand Pathologic Movements? A Pilot Study on Seizure Semiology
Lina Zhang, Tonmoy Monsoor, Mehmet Efe Lorasdagi +8
Multimodal Large Language Models (MLLMs) have demonstrated robust capabilities in recognizing everyday human activities, yet their potential for analyzing clinically significant in…
cs.CV2025
SEER-VAR: Semantic Egocentric Environment Reasoner for Vehicle Augmented Reality
Yuzhi Lai, Shenghai Yuan, Peizheng Li +2
We present SEER-VAR, a novel framework for egocentric vehicle-based augmented reality (AR) that unifies semantic decomposition, Context-Aware SLAM Branches (CASB), and LLM-driven r…