4 papers
PerCache: Predictive Hierarchical Cache for RAG Applications on Mobile Devices
Kaiwei Liu, Liekang Zeng, Lilin Xu +2
Retrieval-augmented generation (RAG) has been extensively used as a de facto paradigm in various large language model (LLM)-driven applications on mobile devices, such as mobile as…
Synera: Synergistic LLM Serving across Device and Cloud at Scale
Genglin Wang, Liekang Zeng, Bufang Yang +6
Large Language Models (LLMs) are becoming key components in various mobile operating systems, driving smart applications like interactive chatbots and personal assistants. While br…
An LLM-Empowered Low-Resolution Vision System for On-Device Human Behavior Understanding
Siyang Jiang, Bufang Yang, Lilin Xu +8
The rapid advancements in Large Vision Language Models (LVLMs) offer the potential to surpass conventional labeling by generating richer, more detailed descriptions of on-device hu…
ContextAgent: Context-Aware Proactive LLM Agents with Open-World Sensory Perceptions
Bufang Yang, Lilin Xu, Liekang Zeng +7
Recent advances in Large Language Models (LLMs) have propelled intelligent agents from reactive responses to proactive support. While promising, existing proactive agents either re…