3 papers
cs.CL2026
Multi-Stage Verification-Centric Framework for Mitigating Hallucination in Multi-Modal RAG
Baiyu Chen, Wilson Wongso, Xiaoqian Hu +2
This paper presents the technical solution developed by team CRUISE for the KDD Cup 2025 Meta Comprehensive RAG Benchmark for Multi-modal, Multi-turn (CRAG-MM) challenge. The chall…
cs.CV2025
A Deep Learning Approach for Augmenting Perceptional Understanding of Histopathology Images
Xiaoqian Hu
In Recent Years, Digital Technologies Have Made Significant Strides In Augmenting-Human-Health, Cognition, And Perception, Particularly Within The Field Of Computational-Pathology.…
cs.CV2025
Bisecle: Binding and Separation in Continual Learning for Video Language Understanding
Yue Tan, Xiaoqian Hu, Hao Xue +2
Frontier vision-language models (VLMs) have made remarkable improvements in video understanding tasks. However, real-world videos typically exist as continuously evolving data stre…