3 papers
cs.CV2025
MapTrace: Scalable Data Generation for Route Tracing on Maps
Artemis Panagopoulou, Aveek Purohit, Achin Kulshrestha +2
While Multimodal Large Language Models have achieved human-like performance on many visual and textual reasoning tasks, their proficiency in fine-grained spatial understanding, suc…
cs.CV2025
EgoSocial: Benchmarking Proactive Intervention Ability of Omnimodal LLMs via Egocentric Social Interaction Perception
Xijun Wang, Tanay Sharma, Achin Kulshrestha +3
As AR/VR technologies become integral to daily life, there's a growing need for AI that understands human social dynamics from an egocentric perspective. However, current LLMs ofte…
cs.HC2025
Geometry Aware Passthrough Mitigates Cybersickness
Trishia El Chemaly, Mohit Goyal, Tinglin Duan +8
Virtual Reality headsets isolate users from the real-world by restricting their perception to the virtual-world. Video See-Through (VST) headsets address this by utilizing world-fa…