Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
A Use-Case Specific Dataset for Measuring Dimensions of Responsible Performance in LLM-generated Text
Alicia Sagae, Chia-Jung Lee, Sandeep Avula +2
Current methods for evaluating large language models (LLMs) typically focus on high-level tasks such as text generation, without targeting a particular AI application. This approac…
cs.CL2025★ 4 cited
Understanding the Repeat Curse in Large Language Models from a Feature Perspective
Junchi Yao, Shu Yang, Jianhua Xu +3
Large language models (LLMs) have made remarkable progress in various domains, yet they often suffer from repetitive text generation, a phenomenon we refer to as the "Repeat Curse"…