Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Diversity-Oriented Fine-Tuning for Uncertainty-Based Hallucination Detection
Qiuyuan Li, Hongliang Dai, Piji Li
Existing hallucination detection methods are typically conducted at the inference stage, without making any modifications to the model itself. In this paper, we are interested in e…
cs.AI2026
Investigating Advanced Reasoning of Large Language Models via Black-Box Environment Interaction
Congchi Yin, Tianyi Wu, Yankai Shu +5
Existing tasks fall short in evaluating reasoning ability of Large Language Models (LLMs) in an interactive, unknown environment. This deficiency leads to the isolated assessment o…