4 papers
Sanity Checks for Long-Form Hallucination Detection
Geigh Zollicoffer, Minh Vu, Hongli Zhan +2
Hallucination detection methods for large language models increasingly operate on chain-of-thought reasoning traces, yet it remains unclear whether they evaluate the reasoning itse…
Discourse Diversity in Multi-Turn Empathic Dialogue
Hongli Zhan, Emma S. Gueorguieva, Javier Hernandez +3
Large language models (LLMs) produce responses rated as highly empathic in single-turn settings (Ayers et al., 2023; Lee et al., 2024), yet they are also known to be formulaic gene…
AI generates well-liked but templatic empathic responses
Emma S. Gueorguieva, Hongli Zhan, Jina Suh +4
Recent research shows that greater numbers of people are turning to Large Language Models (LLMs) for emotional support, and that people rate LLM responses as more empathic than hum…
SPRI: Aligning Large Language Models with Context-Situated Principles
Hongli Zhan, Muneeza Azmat, Raya Horesh +2
Aligning Large Language Models to integrate and reflect human values, especially for tasks that demand intricate human oversight, is arduous since it is resource-intensive and time…