3 papers
cs.CL2026
LH-Deception: Simulating and Understanding LLM Deceptive Behaviors in Long-Horizon Interactions
Yang Xu, Xuanming Zhang, Samuel Yeh +4
Deception is a pervasive feature of human communication and an emerging concern in large language models (LLMs). While recent studies document instances of LLM deception, most eval…
cs.CL2026
Enhancing LLM-Based Data Annotation with Error Decomposition
Zhen Xu, Vedant Khatri, Yijun Dai +4
Large language models offer a scalable alternative to human coding for data annotation tasks, enabling the scale-up of research across data-intensive domains. While LLMs are alread…
cs.CL2025
Generalization or Memorization: Dynamic Decoding for Mode Steering
Xuanming Zhang
Large Language Models (LLMs) exhibit a troubling duality, capable of both remarkable generalization and brittle, verbatim memorization of their training data. This unpredictability…