Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Look Carefully: Adaptive Visual Reinforcements in Multimodal Large Language Models for Hallucination Mitigation
Xingyu Zhu, Kesen Zhao, Liang Yi +4
Multimodal large language models (MLLMs) have achieved remarkable progress in vision-language reasoning, yet they remain vulnerable to hallucination, where generated content deviat…
cs.CV2023
Inter-frame Accelerate Attack against Video Interpolation Models
Junpei Liao, Zhikai Chen, Liang Yi +3
Deep learning based video frame interpolation (VIF) method, aiming to synthesis the intermediate frames to enhance video quality, have been highly developed in the past few years.…