activity
20242026
collaborators

6 papers

cs.RO2026

Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation

He Kong, Zengjue Chen, Qi Wang +6

Vision-language-action (VLA) models have demonstrated remarkable capabilities in robotic manipulation by leveraging pretrained vision-language models. However, existing post-traini…

cs.CV2026

From Bounding Boxes to Visual Reasoning: An On-Policy Data Annotation Tool for Vision-Language Models

Like Zhang, Runliang Niu, Shiqi Wang +5

Vision-language models (VLMs) are rapidly advancing toward sophisticated grounded structured visual reasoning. Training models for such advanced capabilities demands a new genre of…

cs.CV2026

Beyond Retraining: Training-Free Unknown Class Filtering for Source-Free Open Set Domain Adaptation of Vision-Language Models

Yongguang Li, Jindong Li, Qi Wang +4

Vision-language models (VLMs) have gained widespread attention for their strong zero-shot capabilities across numerous downstream tasks. However, these models assume that each test…

cs.LG2025

GLADMamba: Unsupervised Graph-Level Anomaly Detection Powered by Selective State Space Model

Yali Fu, Jindong Li, Qi Wang +1

Unsupervised graph-level anomaly detection (UGLAD) is a critical and challenging task across various domains, such as social network analysis, anti-cancer drug discovery, and toxic…

cs.LG2024

HC-GLAD: Dual Hyperbolic Contrastive Learning for Unsupervised Graph-Level Anomaly Detection

Yali Fu, Jindong Li, Jiahong Liu +3

Unsupervised graph-level anomaly detection (UGAD) has garnered increasing attention in recent years due to its significance. Most existing methods that rely on traditional GNNs mai…

cs.IR2024

Towards Next-Generation LLM-based Recommender Systems: A Survey and Beyond

Qi Wang, Jindong Li, Shiqi Wang +7

Large language models (LLMs) have not only revolutionized the field of natural language processing (NLP) but also have the potential to bring a paradigm shift in many other fields…