Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Principled Detection of Hallucinations in Large Language Models via Multiple Testing
Jiawei Li, Akshayaa Magesh, Venugopal V. Veeravalli
While Large Language Models (LLMs) have emerged as powerful foundational models to solve a variety of tasks, they have also been shown to be prone to hallucinations, i.e., generati…
cs.CL2026
Towards Trustworthy Multimodal Moderation via Policy-Aligned Reasoning and Hierarchical Labeling
Anqi Li, Wenwei Jin, Jintao Tong +3
Social platforms have revolutionized information sharing, but also accelerated the dissemination of harmful and policy-violating content. To ensure safety and compliance at scale,…