Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
RP-OPSD: Resolution-Privileged On-Policy Self-Distillation for Multimodal Large Language Models
Qihui Zhu, Yuchen Wang, Zijian Wen +7
On-Policy Self-Distillation (OPSD) uses privileged information available only to the teacher to provide dense token-level supervision on trajectories generated by the student. Howe…
cs.CV2024
LIME: Less Is More for MLLM Evaluation
King Zhu, Qianbo Zang, Shian Jia +18
Multimodal Large Language Models (MLLMs) are evaluated on various benchmarks, such as image captioning, visual question answering, and reasoning. However, many of these benchmarks…