collaborators

10 papers

cs.LG2026

Self-Evolving Multi-Agent Systems via Textual Backpropagation

Xiaowen Ma, Yunpu Ma, Chenyang Lin +6

Leveraging multiple Large Language Models (LLMs) has proven effective for addressing complex, high-dimensional tasks, but current approaches often rely on static, manually engineer…

cs.CV2026

PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection

Jinhe Bi, Aniri, Zengjie Jin +11

Visual instruction tuning adapts pre-trained Multimodal Large Language Models (MLLMs) to follow human instructions for real-world applications. However, the rapid growth of these d…

quant-ph2026

Quantum Architecture Search with Unsupervised Representation Learning

Yize Sun, Zixin Wu, Volker Tresp +1

Unsupervised representation learning presents new opportunities for advancing Quantum Architecture Search (QAS) on Noisy Intermediate-Scale Quantum (NISQ) devices. QAS is designed…

cs.CV2026

Multi-event Video-Text Retrieval

Gengyuan Zhang, Jisen Ren, Jindong Gu +1

Video-Text Retrieval (VTR) is a crucial multi-modal task in an era of massive video-text data on the Internet. A plethora of work characterized by using a two-stream Vision-Languag…

cs.CV2025

METok: Multi-Stage Event-based Token Compression for Efficient Long Video Understanding

Mengyue Wang, Shuo Chen, Kristian Kersting +2

Recent advances in Video Large Language Models (VLLMs) have significantly enhanced their ability to understand video content. Nonetheless, processing long videos remains challengin…

cs.AI2025

SwarmAgentic: Towards Fully Automated Agentic System Generation via Swarm Intelligence

Yao Zhang, Chenyang Lin, Shijie Tang +4

The rapid progress of Large Language Models has advanced agentic systems in decision-making, coordination, and task execution. Yet, existing agentic system generation frameworks la…