2 papers
cs.CL2026
OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Scenarios
Xinyi Li, Zhen Fang, Yongxin Deng +12
Hallucination detection is essential for the reliable deployment of large language models (LLMs). However, existing evaluations face two core challenges: inconsistent inference con…
cs.AI2025
Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations
Jinyuan Luo, Zhen Fang, Yixuan Li +2
Hallucination remains a key obstacle to the reliable deployment of large language models (LLMs) in real-world question answering tasks. A widely adopted strategy to detect hallucin…