2 papers
cs.CV2025
3D Prior is All You Need: Cross-Task Few-shot 2D Gaze Estimation
Yihua Cheng, Hengfei Wang, Zhongqun Zhang +4
3D and 2D gaze estimation share the fundamental objective of capturing eye movements but are traditionally treated as two distinct research domains. In this paper, we introduce a n…
cs.CV2025
Hierarchical Context Transformer for Multi-level Semantic Scene Understanding
Luoying Hao, Yan Hu, Yang Yue +4
A comprehensive and explicit understanding of surgical scenes plays a vital role in developing context-aware computer-assisted systems in the operating theatre. However, few works…