2 papers
cs.SE2026
ToolCaching: Towards Efficient Caching for LLM Tool-calling
Yi Zhai, Dian Shen, Junzhou Luo +1
Recent advances in Large Language Models (LLMs) have revolutionized web applications, enabling intelligent search, recommendation, and assistant services with natural language inte…
cs.MM2025
EV-NVC: Efficient Variable bitrate Neural Video Compression
Yongcun Hu, Yingzhen Zhai, Jixiang Luo +4
Training neural video codec (NVC) with variable rate is a highly challenging task due to its complex training strategies and model structure. In this paper, we train an efficient v…