2 papers
cs.AI2024
How Different AI Chatbots Behave? Benchmarking Large Language Models in Behavioral Economics Games
Yutong Xie, Yiyao Liu, Zhuang Ma +5
The deployment of large language models (LLMs) in diverse applications requires a thorough understanding of their decision-making strategies and behavioral patterns. As a supplemen…
cs.CL2024
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG
Xinxi Chen, Li Wang, Wei Wu +2
Hallucination is a key roadblock for applications of Large Language Models (LLMs), particularly for enterprise applications that are sensitive to information accuracy. To address t…