6 citations · 7 across the 6 of their papers we have counts for
8 papers
DeRelayL: Sustainable Decentralized Relay Learning
Haihan Duan, Tengfei Ma, Yuyang Qin +4
In the era of big data, large-scale machine learning models have revolutionized various fields, driving significant advancements. However, large-scale model training demands high f…
Sparse Shortcuts: Facilitating Efficient Fusion in Multimodal Large Language Models
Jingrui Zhang, Feng Liang, Yong Zhang +3
With the remarkable success of large language models (LLMs) in natural language understanding and generation, multimodal large language models (MLLMs) have rapidly advanced in thei…
Revisiting Cross-Architecture Distillation: Adaptive Dual-Teacher Transfer for Lightweight Video Models
Ying Peng, Hongsen Ye, Changxin Huang +3
Vision Transformers (ViTs) have achieved strong performance in video action recognition, but their high computational cost limits their practicality. Lightweight CNNs are more effi…
CO-PFL: Contribution-Oriented Personalized Federated Learning for Heterogeneous Networks
Ke Xing, Yanjie Dong, Xiaoyi Fan +4
Personalized federated learning (PFL) addresses a critical challenge of collaboratively training customized models for clients with heterogeneous and scarce local data. Conventiona…
OVG-HQ: Online Video Grounding with Hybrid-modal Queries
Runhao Zeng, Jiaqi Mao, Minghao Lai +5
Video grounding (VG) task focuses on locating specific moments in a video based on a query, usually in text form. However, traditional VG struggles with some scenarios like streami…
Emotion Recognition from Skeleton Data: A Comprehensive Survey
Haifeng Lu, Jiuyi Chen, Zhen Zhang +3
Emotion recognition through body movements has emerged as a compelling and privacy-preserving alternative to traditional methods that rely on facial expressions or physiological si…