1 paper
Zijun Liu, Boqun Kou, Peng Li +4
Despite the strong performance of large language models (LLMs) across a wide range of tasks, they still have reliability issues. Previous studies indicate that strong LLMs like GPT…