1 citations · 2 across the 5 of their papers we have counts for
7 papers
PerCache: Predictive Hierarchical Cache for RAG Applications on Mobile Devices
Kaiwei Liu, Liekang Zeng, Lilin Xu +2
Retrieval-augmented generation (RAG) has been extensively used as a de facto paradigm in various large language model (LLM)-driven applications on mobile devices, such as mobile as…
A Large-Scale Multimodal Dataset and Benchmarks for Human Activity Scene Understanding and Reasoning
Siyang Jiang, Mu Yuan, Xiang Ji +12
Multimodal human action recognition (HAR) leverages complementary sensors for activity classification. Beyond recognition, recent advances in large language models (LLMs) enable de…
An LLM-Empowered Low-Resolution Vision System for On-Device Human Behavior Understanding
Siyang Jiang, Bufang Yang, Lilin Xu +8
The rapid advancements in Large Vision Language Models (LVLMs) offer the potential to surpass conventional labeling by generating richer, more detailed descriptions of on-device hu…
ContextAgent: Context-Aware Proactive LLM Agents with Open-World Sensory Perceptions
Bufang Yang, Lilin Xu, Liekang Zeng +7
Recent advances in Large Language Models (LLMs) have propelled intelligent agents from reactive responses to proactive support. While promising, existing proactive agents either re…
Exploring the Capabilities of LLMs for IMU-based Fine-grained Human Activity Understanding
Lilin Xu, Kaiyuan Hou, Xiaofan Jiang
Human activity recognition (HAR) using inertial measurement units (IMUs) increasingly leverages large language models (LLMs), yet existing approaches focus on coarse activities lik…
TDBench: A Benchmark for Top-Down Image Understanding with Reliability Analysis of Vision-Language Models
Kaiyuan Hou, Minghui Zhao, Lilin Xu +2
Top-down images play an important role in safety-critical settings such as autonomous navigation and aerial surveillance, where they provide holistic spatial information that front…