1 citations · 1 across the 4 of their papers we have counts for
Showing cs.DCShow all
2 papers · 1 filter
cs.DC2025
PerCache: Predictive Hierarchical Cache for RAG Applications on Mobile Devices
Kaiwei Liu, Liekang Zeng, Lilin Xu +2
Retrieval-augmented generation (RAG) has been extensively used as a de facto paradigm in various large language model (LLM)-driven applications on mobile devices, such as mobile as…
cs.DC2025
Synera: Synergistic LLM Serving across Device and Cloud at Scale
Genglin Wang, Liekang Zeng, Bufang Yang +6
Large Language Models (LLMs) are becoming key components in various mobile operating systems, driving smart applications like interactive chatbots and personal assistants. While br…