collaborators

10 papers

cs.CV2025

Language-Instructed Reasoning for Group Activity Detection via Multimodal Large Language Model

Jihua Peng, Qianxiong Xu, Yichen Liu +4

Group activity detection (GAD) aims to simultaneously identify group members and categorize their collective activities within video sequences. Existing deep learning-based methods…

cs.CV2025

SAMITE: Position Prompted SAM2 with Calibrated Memory for Visual Object Tracking

Qianxiong Xu, Lanyun Zhu, Chenxi Liu +4

Visual Object Tracking (VOT) is widely used in applications like autonomous driving to continuously track targets in videos. Existing methods can be roughly categorized into templa…

cs.LG2025

LLMs Meet Cross-Modal Time Series Analytics: Overview and Directions

Chenxi Liu, Hao Miao, Cheng Long +3

Large Language Models (LLMs) have emerged as a promising paradigm for time series analytics, leveraging their massive parameters and the shared sequential nature of textual and tim…

cs.CV2025

Unlocking the Power of SAM 2 for Few-Shot Segmentation

Qianxiong Xu, Lanyun Zhu, Xuanyi Liu +4

Few-Shot Segmentation (FSS) aims to learn class-agnostic segmentation on few classes to segment arbitrary classes, but at the risk of overfitting. To address this, some methods use…

cs.LG2025

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache

Qiuyu Zhu, Liang Zhang, Qianxiong Xu +2

Graph-based retrieval-augmented generation (RAG) enables large language models (LLMs) to incorporate structured knowledge via graph retrieval as contextual input, enhancing more ac…

cs.LG2025

Efficient Multivariate Time Series Forecasting via Calibrated Language Models with Privileged Knowledge Distillation

Chenxi Liu, Hao Miao, Qianxiong Xu +5

Multivariate time series forecasting (MTSF) endeavors to predict future observations given historical data, playing a crucial role in time series data management systems. With adva…