3 papers
cs.CV2026
ClickAIXR: On-Device Multimodal Vision-Language Interaction with Real-World Objects in Extended Reality
Dawar Khan, Alexandre Kouyoumdjian, Xinyu Liu +3
We present ClickAIXR, a novel on-device framework for multimodal vision-language interaction with objects in extended reality (XR). Unlike prior systems that rely on cloud-based AI…
cs.DC2025
AIvaluateXR: An Evaluation Framework for on-Device AI in XR with Benchmarking Results
Dawar Khan, Xinyu Liu, Omar Mena +3
The deployment of large language models (LLMs) on extended reality (XR) devices has great potential to advance the field of human-AI interaction. In the case of direct, on-device m…
cs.HC2025
Augmenting a Large Language Model with a Combination of Text and Visual Data for Conversational Visualization of Global Geospatial Data
Omar Mena, Alexandre Kouyoumdjian, Lonni Besançon +3
We present a method for augmenting a Large Language Model (LLM) with a combination of text and visual data to enable accurate question answering in visualization of scientific data…