Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
M2-Verify: A Large-Scale Multidomain Benchmark for Checking Multimodal Claim Consistency
Abolfazl Ansari, Delvin Ce Zhang, Zhuoyang Zou +2
Evaluating scientific arguments requires assessing the strict consistency between a claim and its underlying multimodal evidence. However, existing benchmarks lack the scale, domai…
cs.CL2026
Echoes of Automation: The Increasing Use of LLMs in Newsmaking
Abolfazl Ansari, Delvin Ce Zhang, Nafis Irtiza Tripto +1
The rapid rise of Generative AI (GenAI), particularly LLMs, poses concerns for journalistic integrity and authorship. This study examines AI-generated content across over 40,000 ne…