75 citations · 80 across the 24 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
VisualQuest: A Benchmark for Abstract Visual Reasoning in MLLMs
Kelaiti Xiao, Liang Yang, Dongyu Zhang +2
We introduce VisualQuest, a novel dataset designed to rigorously evaluate multimodal large language models (MLLMs) on abstract visual reasoning tasks that require the integration o…
cs.CV2024
Towards Patronizing and Condescending Language in Chinese Videos: A Multimodal Dataset and Detector
Hongbo Wang, Junyu Lu, Yan Han +3
Patronizing and Condescending Language (PCL) is a form of discriminatory toxic speech targeting vulnerable groups, threatening both online and offline safety. While toxic speech re…