3 papers
cs.CL2025
Curse of High Dimensionality Issue in Transformer for Long-context Modeling
Shuhai Zhang, Zeng You, Yaofo Chen +5
Transformer-based large language models (LLMs) excel in natural language processing tasks by capturing long-range dependencies through self-attention mechanisms. However, long-cont…
cs.CV2025
Zero-Shot Skeleton-Based Action Recognition With Prototype-Guided Feature Alignment
Kai Zhou, Shuhai Zhang, Zeng You +3
Zero-shot skeleton-based action recognition aims to classify unseen skeleton-based human actions without prior exposure to such categories during training. This task is extremely c…
cs.CV2024
Towards Long Video Understanding via Fine-detailed Video Story Generation
Zeng You, Zhiquan Wen, Yaofo Chen +4
Long video understanding has become a critical task in computer vision, driving advancements across numerous applications from surveillance to content retrieval. Existing video und…