Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Quantitative LLM Judges
Aishwarya Sahoo, Jeevana Kruthi Karnuthala, Tushar Parmanand Budhwani +9
LLM-as-a-judge is a framework where a large language model (LLM) evaluates the output of another LLM. While LLMs excel at producing qualitative textual evaluations, they often stru…
cs.CL2025
A Penalty Goes a Long Way: Measuring Lexical Diversity in Synthetic Texts Under Prompt-Influenced Length Variations
Vijeta Deshpande, Ishita Dasgupta, Uttaran Bhattacharya +3
Synthetic text generated by Large Language Models (LLMs) is increasingly used for further training and improvement of LLMs. Diversity is crucial for the effectiveness of synthetic…
cs.CL2025
Detecting Ambiguities to Guide Query Rewrite for Robust Conversations in Enterprise AI Assistants
Md Mehrab Tanjim, Xiang Chen, Victor S. Bursztyn +8
Multi-turn conversations with an Enterprise AI Assistant can be challenging due to conversational dependencies in questions, leading to ambiguities and errors. To address this, we…