Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs
Xiabin Zhou, Wenbin Wang, Minyan Zeng +5
Efficient KV cache management in LLMs is crucial for long-context tasks like RAG and summarization. Existing KV cache compression methods enforce a fixed pattern, neglecting task-s…
cs.CL2024
Towards Robust Multimodal Sentiment Analysis with Incomplete Data
Haoyu Zhang, Wenbin Wang, Tianshu Yu
The field of Multimodal Sentiment Analysis (MSA) has recently witnessed an emerging direction seeking to tackle the issue of data incompleteness. Recognizing that the language moda…