2 papers
cs.CV2025
MetaSlot: Break Through the Fixed Number of Slots in Object-Centric Learning
Hongjia Liu, Rongzhen Zhao, Haohan Chen +1
Learning object-level, structured representations is widely regarded as a key to better generalization in vision and underpins the design of next-generation Pre-trained Vision Mode…
cs.CV2025
DMAGaze: Gaze Estimation Based on Feature Disentanglement and Multi-Scale Attention
Haohan Chen, Hongjia Liu, Shiyong Lan +4
Gaze estimation, which predicts gaze direction, commonly faces the challenge of interference from complex gaze-irrelevant information in face images. In this work, we propose DMAGa…