Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness
Hao An, Yibin Lou, Jiayi Guo +1
Large language models (LLMs) often exhibit hallucinations due to their inability to accurately perceive their own knowledge boundaries. Existing abstention fine-tuning methods typi…
cs.CL2025
Teaching LLMs to Abstain via Fine-Grained Semantic Confidence Reward
Hao An, Yang Xu
Mitigating hallucinations in Large Language Models (LLMs) is critical for their reliable deployment. Existing methods typically fine-tune LLMs to abstain from answering questions b…
cs.CL2024
Detecting Subtle Differences between Human and Model Languages Using Spectrum of Relative Likelihood
Yang Xu, Yu Wang, Hao An +2
Human and model-generated texts can be distinguished by examining the magnitude of likelihood in language. However, it is becoming increasingly difficult as language model's capabi…