4 citations · 4 across the 4 of their papers we have counts for
4 papers
Motion-Based Tokenization for Cross-Dataset Egocentric Gaze Modeling
Virmarie Maquiling, Zhuojiang Cai, Enkelejda Kasneci
Gaze is increasingly used as an input signal for vision and multimodal models, yet no consensus exists on how to represent it across datasets. Raw traces preserve detail but are no…
GazeOnce360: Fisheye-Based 360° Multi-Person Gaze Estimation with Global-Local Feature Fusion
Zhuojiang Cai, Zhenghui Sun, Feng Lu
We present GazeOnce360, a novel end-to-end model for multi-person gaze estimation from a single tabletop-mounted upward-facing fisheye camera. Unlike conventional approaches that r…
GazeSwipe: Enhancing Mobile Touchscreen Reachability through Seamless Gaze and Finger-Swipe Integration
Zhuojiang Cai, Jingkai Hong, Zhimin Wang +1
Smartphones with large screens provide users with increased display and interaction space but pose challenges in reaching certain areas with the thumb when using the device with on…
Robust Dual-Modal Speech Keyword Spotting for XR Headsets
Zhuojiang Cai, Yuhan Ma, Feng Lu
While speech interaction finds widespread utility within the Extended Reality (XR) domain, conventional vocal speech keyword spotting systems continue to grapple with formidable ch…