1 paper · 1 filter
Cheng Gao, Cheng Huang, Kangyang Luo +5
Enabling large language models (LLMs) to appropriately abstain from answering questions beyond their knowledge is crucial for mitigating hallucinations. While existing reinforcemen…