7 papers
GMO-EDIT: Grounded Multi-Operation Editing for E-Commerce Images
Zipeng Guo, Xiaoan Liu, Lichen Ma +9
Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature poses a dual challenge: mod…
VisionClaw: Always-On AI Agents through Smart Glasses
Xiaoan Liu, DaeHo Lee, Eric J Gonzalez +2
We present VisionClaw, an always-on wearable AI agent that integrates live egocentric perception with agentic task execution. Running on Meta Ray-Ban smart glasses, VisionClaw cont…
Semantic Reality: Interactive Context-Aware Visualization of Inter-Object Relationships in Augmented Reality
Xiaoan Liu, Eric J Gonzalez, Nels Numan +5
Bridging the physical and digital world through interaction remains a core challenge in augmented reality (AR). Existing systems target single objects, limiting support for plannin…
Conversational Successes and Breakdowns in Everyday Smart Glasses Use
Xiuqi Tommy Zhu, Xiaoan Liu, Casper Harteveld +2
Non-Display Smart Glasses hold the potential to support everyday activities by combining continuous environmental sensing with voice-only interaction powered by large language mode…
Generative Lecture: Making Lecture Videos Interactive with LLMs and AI Clone Instructors
Hye-Young Jo, Ada Yi Zhao, Xiaoan Liu +1
We introduce Generative Lecture, a concept that makes existing lecture videos interactive through generative AI and AI clone instructors. By leveraging interactive avatars powered…
Reality Proxy: Fluid Interactions with Real-World Objects in MR via Abstract Representations
Xiaoan Liu, Difan Jia, Xianhao Carton Liu +2
Interacting with real-world objects in Mixed Reality (MR) often proves difficult when they are crowded, distant, or partially occluded, hindering straightforward selection and mani…