2 papers
cs.CV2025
Distinguishing Visually Similar Actions: Prompt-Guided Semantic Prototype Modulation for Few-Shot Action Recognition
Xiaoyang Li, Mingming Lu, Ruiqi Wang +2
Few-shot action recognition aims to enable models to quickly learn new action categories from limited labeled samples, addressing the challenge of data scarcity in real-world appli…
cs.AI2025
Efficiently Enhancing General Agents With Hierarchical-categorical Memory
Changze Qiao, Mingming Lu
With large language models (LLMs) demonstrating remarkable capabilities, there has been a surge in research on leveraging LLMs to build general-purpose multi-modal agents. However,…