6 papers
MISID: A Multimodal Multi-turn Dataset for Complex Intent Recognition in Strategic Deception Games
Shufang Lin, Muyang Chen, Xiabing Zhou +3
Understanding human intent in complex multi-turn interactions remains a fundamental challenge in human-computer interaction and behavioral analysis. While existing intent recogniti…
RepCaM++: Exploring Transparent Visual Prompt With Inference-Time Re-Parameterization for Neural Video Delivery
Rongyu Zhang, Xize Duan, Jiaming Liu +5
Recently, content-aware methods have been employed to reduce bandwidth and enhance the quality of Internet video delivery. These methods involve training distinct content-aware sup…
Fine-Tuning and Deploying Large Language Models Over Edges: Issues and Approaches
Yanjie Dong, Haijun Zhang, Chengming Li +3
Since the release of GPT2-1.5B in 2019, the large language models (LLMs) have evolved from specialized deep models to versatile foundation models. While demonstrating remarkable ze…
DeformStream: Deformation-based Adaptive Volumetric Video Streaming
Boyan Li, Yongting Chen, Dayou Zhang +1
Volumetric video streaming offers immersive 3D experiences but faces significant challenges due to high bandwidth requirements and latency issues in transmitting detailed content i…
HeadsetOff: Enabling Photorealistic Video Conferencing on Economical VR Headsets
Yili Jin, Xize Duan, Fangxin Wang +1
Virtual Reality (VR) has become increasingly popular for remote collaboration, but video conferencing poses challenges when the user's face is covered by the headset. Existing solu…
Multi-level Personalized Federated Learning on Heterogeneous and Long-Tailed Data
Rongyu Zhang, Yun Chen, Chenrui Wu +2
Federated learning (FL) offers a privacy-centric distributed learning framework, enabling model training on individual clients and central aggregation without necessitating data ex…