4 papers
Diving into Mitigating Hallucinations from a Vision Perspective for Large Vision-Language Models
Weihang Wang, Xinhao Li, Ziyue Wang +5
Object hallucination in Large Vision-Language Models (LVLMs) significantly impedes their real-world applicability. As the primary component for accurately interpreting visual infor…
Data Valuation and Selection in a Federated Model Marketplace
Wenqian Li, Youjia Yang, Ruoxi Jia +1
In the era of Artificial Intelligence (AI), marketplaces have become essential platforms for facilitating the exchange of data products to foster data sharing. Model transactions p…
Commuting Distance Regularization for Timescale-Dependent Label Inconsistency in EEG Emotion Recognition
Xiaocong Zeng, Craig Michoski, Yan Pang +1
In this work, we address the often-overlooked issue of Timescale Dependent Label Inconsistency (TsDLI) in training neural network models for EEG-based human emotion recognition. To…
Multimodal Misinformation Detection by Learning from Synthetic Data with Multimodal LLMs
Fengzhu Zeng, Wenqian Li, Wei Gao +1
Detecting multimodal misinformation, especially in the form of image-text pairs, is crucial. Obtaining large-scale, high-quality real-world fact-checking datasets for training dete…