3 papers
cs.CV2025
EventHallusion: Diagnosing Event Hallucinations in Video LLMs
Jiacheng Zhang, Yang Jiao, Shaoxiang Chen +5
Recently, Multimodal Large Language Models (MLLMs) have made significant progress in the video comprehension field. Despite remarkable content reasoning and instruction following c…
cs.MM2025
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
Junhao Xu, Jingjing Chen, Yang Jiao +4
Recent advances in generative artificial intelligence have enabled the creation of highly realistic image forgeries, raising significant concerns about digital media authenticity.…
cs.CV2024
EAGLE: Towards Efficient Arbitrary Referring Visual Prompts Comprehension for Multimodal Large Language Models
Jiacheng Zhang, Yang Jiao, Shaoxiang Chen +2
Recently, Multimodal Large Language Models (MLLMs) have sparked great research interests owing to their exceptional content-reasoning and instruction-following capabilities. To eff…