2 papers
cs.CL2026
OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models
Jianjiang Yang, Peihang Li, Shanqing Xu +3
While Multimodal Large Language Models (MLLMs) have achieved remarkable progress across diverse tasks, they suffer from hallucinations where generated outputs contradict or misrepr…
cs.CV2026
RIDGE: Region-Informed Derivative-Guided Evidence Selection for Long Video Understanding
Shanqing Xu, Meng Luo, Mengchen Qian +7
Long videos contain far more visual content than Large Vision-Language Models (LVLMs) can process under a fixed visual-token budget, making frame selection essential. Existing quer…