3 papers
cs.CV2025
Exploring the Capabilities of LLMs for IMU-based Fine-grained Human Activity Understanding
Lilin Xu, Kaiyuan Hou, Xiaofan Jiang
Human activity recognition (HAR) using inertial measurement units (IMUs) increasingly leverages large language models (LLMs), yet existing approaches focus on coarse activities lik…
cs.LG2025
TDBench: A Benchmark for Top-Down Image Understanding with Reliability Analysis of Vision-Language Models
Kaiyuan Hou, Minghui Zhao, Lilin Xu +2
Top-down images play an important role in safety-critical settings such as autonomous navigation and aerial surveillance, where they provide holistic spatial information that front…
cs.HC2024
Visualizing the Invisible: A Generative AR System for Intuitive Multi-Modal Sensor Data Presentation
Yunqi Guo, Kaiyuan Hou, Heming Fu +4
Understanding sensor data can be difficult for non-experts because of the complexity and different semantic meanings of sensor modalities. This leads to a need for intuitive and ef…