14 papers
Decoding Multimodal Cues: Unveiling the Implicit Meaning Behind Hateful Videos
Junyu Lu, Deyi Ji, Liqun Liu +9
Hateful videos have become prevalent on online platforms, highlighting an urgent need for effective detection. However, existing studies primarily focus on binary classification an…
PersonaAgent: Bridging Memory and Action for Personalized LLM Agents
Weizhi Zhang, Xinyang Zhang, Chenwei Zhang +12
Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of domains and tasks. Despite their pot…
Distinguishing Right from Wrong in Debates: Attribution Analysis of Chinese Harmful Memes
Weiming Wang, Junyu Lu, Han Wang +5
Research on harmful meme detection has garnered significant attention, resulting in the development of numerous datasets and methods. However, progress in detecting Chinese harmful…
Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis
Junyu Lu, Deyi Ji, Xuanyi Liu +5
Large language models for subjectivity analysis are typically trained with aggregated labels, which compress variations in human judgment into a single supervision signal. This par…
Propagating Similarity, Mitigating Uncertainty: Similarity Propagation-enhanced Uncertainty for Multimodal Recommendation
Xinzhuo Wu, Hongbo Wang, Yuan Lin +3
Multimodal Recommendation (MMR) systems are crucial for modern platforms but are often hampered by inherent noise and uncertainty in modal features, such as blurry images, diverse…
VisualQuest: A Benchmark for Abstract Visual Reasoning in MLLMs
Kelaiti Xiao, Liang Yang, Dongyu Zhang +2
We introduce VisualQuest, a novel dataset designed to rigorously evaluate multimodal large language models (MLLMs) on abstract visual reasoning tasks that require the integration o…