1 paper · 1 filter
Hanyu Wang, Bochuan Cao, Yuanpu Cao +1
Large language models (LLMs) are known to struggle with consistently generating truthful responses. While various representation intervention techniques have been proposed, these m…