5 papers · 1 filter
MANZANO: A Simple and Scalable Unified Multimodal Model with a Hybrid Vision Tokenizer
Yanghao Li, Rui Qian, Bowen Pan +24
Unified multimodal Large Language Models (LLMs) that can both understand and generate visual content hold immense potential. However, existing open-source models often suffer from…
EvHand-FPV: Efficient Event-Based 3D Hand Tracking from First-Person View
Zhen Xu, Guorui Lu, Chang Gao +1
Hand tracking holds great promise for intuitive interaction paradigms, but frame-based methods often struggle to meet the requirements of accuracy, low latency, and energy efficien…
Event-Based Eye Tracking. 2025 Event-based Vision Workshop
Qinyu Chen, Chang Gao, Min Liu +29
This survey serves as a review for the 2025 Event-Based Eye Tracking Challenge organized as part of the 2025 CVPR event-based vision workshop. This challenge focuses on the task of…
SlimSeiz: Efficient Channel-Adaptive Seizure Prediction Using a Mamba-Enhanced Network
Guorui Lu, Jing Peng, Bingyuan Huang +4
Epileptic seizures cause abnormal brain activity, and their unpredictability can lead to accidents, underscoring the need for long-term seizure prediction. Although seizures can be…
FACET: Fast and Accurate Event-Based Eye Tracking Using Ellipse Modeling for Extended Reality
Junyuan Ding, Ziteng Wang, Chang Gao +2
Eye tracking is a key technology for gaze-based interactions in Extended Reality (XR), but traditional frame-based systems struggle to meet XR's demands for high accuracy, low late…