Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
CFPO: Counterfactual Policy Optimization for Multimodal Reasoning
Zhangyuan Yu, Wanran Sun, Guangjing Yang +2
Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal reasoning. However, prevailing reinforcement learning (RL) paradigms lack explicit coun…
cs.CV2025
iDPA: Instance Decoupled Prompt Attention for Incremental Medical Object Detection
Huahui Yi, Wei Xu, Ziyuan Qin +4
Existing prompt-based approaches have demonstrated impressive performance in continual learning, leveraging pre-trained large-scale models for classification tasks; however, the ti…
cs.CV2024
MLAE: Masked LoRA Experts for Visual Parameter-Efficient Fine-Tuning
Junjie Wang, Guangjing Yang, Wentao Chen +4
In response to the challenges posed by the extensive parameter updates required for full fine-tuning of large-scale pre-trained models, parameter-efficient fine-tuning (PEFT) metho…