3 papers
cs.CL2026
Multi-Stage Verification-Centric Framework for Mitigating Hallucination in Multi-Modal RAG
Baiyu Chen, Wilson Wongso, Xiaoqian Hu +2
This paper presents the technical solution developed by team CRUISE for the KDD Cup 2025 Meta Comprehensive RAG Benchmark for Multi-modal, Multi-turn (CRAG-MM) challenge. The chall…
cs.CV2025
Bisecle: Binding and Separation in Continual Learning for Video Language Understanding
Yue Tan, Xiaoqian Hu, Hao Xue +2
Frontier vision-language models (VLMs) have made remarkable improvements in video understanding tasks. However, real-world videos typically exist as continuously evolving data stre…
cs.DB2025
ArcNeural: A Multi-Modal Database for the Gen-AI Era
Wu Min, Qiao Yuncong, Yu Tan +1
ArcNeural introduces a novel multimodal database tailored for the demands of Generative AI and Large Language Models, enabling efficient management of diverse data types such as gr…