1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AR2025★ 1 cited
CLONE: Customizing LLMs for Efficient Latency-Aware Inference at the Edge
Chunlin Tian, Xinpeng Qin, Kahou Tam +5
Deploying large language models (LLMs) on edge devices is crucial for delivering fast responses and ensuring data privacy. However, the limited storage, weight, and power of edge d…
eess.IV2024
Uniformly Accelerated Motion Model for Inter Prediction
Zhuoyuan Li, Yao Li, Chuanbo Tang +3
Inter prediction is a key technology to reduce the temporal redundancy in video coding. In natural videos, there are usually multiple moving objects with variable velocity, resulti…