most citedInfinite Video Understanding

1 citations · 1 across the 3 of their papers we have counts for

collaborators

6 papers

cs.CV2025

TeleEgo: Benchmarking Egocentric AI Assistants in the Wild

Jiaqi Yan, Ruilong Ren, Jingren Liu +13

Egocentric AI assistants in real-world settings must process multi-modal inputs (video, audio, text), respond in real time, and retain evolving long-term memory. However, existing…

cs.CV20251 cited

Infinite Video Understanding

Dell Zhang, Xiangyu Chen, Jixiang Luo +6

The rapid advancements in Large Language Models (LLMs) and their multimodal extensions (MLLMs) have ushered in remarkable progress in video understanding. However, a fundamental ch…

cs.CV2025

Interpretable Few-Shot Image Classification via Prototypical Concept-Guided Mixture of LoRA Experts

Zhong Ji, Rongshuai Wei, Jingren Liu +2

Self-Explainable Models (SEMs) rely on Prototypical Concept Learning (PCL) to enable their visual recognition processes more interpretable, but they often struggle in data-scarce s…

cs.LG2025

CCD: Continual Consistency Diffusion for Lifelong Generative Modeling

Jingren Liu, Shuning Xu, Yun Wang +2

While diffusion-based models have shown remarkable generative capabilities in static settings, their extension to continual learning (CL) scenarios remains fundamentally constraine…

cs.CV2025

Optimal Transport Adapter Tuning for Bridging Modality Gaps in Few-Shot Remote Sensing Scene Classification

Zhong Ji, Ci Liu, Jingren Liu +3

Few-Shot Remote Sensing Scene Classification (FS-RSSC) presents the challenge of classifying remote sensing images with limited labeled samples. Existing methods typically emphasiz…

cs.CV2024

Multi-Stage Knowledge Integration of Vision-Language Models for Continual Learning

Hongsheng Zhang, Zhong Ji, Jingren Liu +2

Vision Language Models (VLMs), pre-trained on large-scale image-text datasets, enable zero-shot predictions for unseen data but may underperform on specific unseen tasks. Continual…