5 papers
Zero-Shot Personalized Camera Motion Control for Image-to-Video Synthesis
Pooja Guhan, Divya Kothandaraman, Geonsun Lee +3
Specifying nuanced and compelling camera motion remains a significant hurdle for non-expert creators using generative tools, creating an "expressive gap" where generic text prompts…
SDS KoPub VDR: A Benchmark Dataset for Visual Document Retrieval in Korean Public Documents
Jaehoon Lee, Sohyun Kim, Wanggeun Park +3
Existing benchmarks for visual document retrieval (VDR) largely overlook non-English languages and the structural complexity of official publications. To address this gap, we intro…
Sensible Agent: A Framework for Unobtrusive Interaction with Proactive AR Agents
Geonsun Lee, Min Xia, Nels Numan +9
Proactive AR agents promise context-aware assistance, but their interactions often rely on explicit voice prompts or responses, which can be disruptive or socially awkward. We intr…
XR Blocks: Accelerating Human-centered AI + XR Innovation
David Li, Nels Numan, Xun Qian +17
We are on the cusp where Artificial Intelligence (AI) and Extended Reality (XR) are converging to unlock new paradigms of interactive computing. However, a significant gap exists b…
Since U Been Gone: Augmenting Context-Aware Transcriptions for Re-engaging in Immersive VR Meetings
Geonsun Lee, Yue Yang, Jennifer Healey +1
Maintaining engagement in immersive meetings is challenging, particularly when users must catch up on missed content after disruptions. While transcription interfaces can help, tab…