2 papers
cs.CL2026
Evaluating and Calibrating LLM Confidence on Questions with Multiple Correct Answers
Yuhan Wang, Shiyu Ni, Zhikai Ding +3
Confidence calibration is essential for making large language models (LLMs) reliable, yet existing training-free methods have been primarily studied under single-answer question an…
cs.CL2025
Do LVLMs Know What They Know? A Systematic Study of Knowledge Boundary Perception in LVLMs
Zhikai Ding, Shiyu Ni, Keping Bi
Large vision-language models (LVLMs) demonstrate strong visual question answering (VQA) capabilities but are shown to hallucinate. A reliable model should perceive its knowledge bo…