ai assistants 1benchmark dataset 1dialogue systems 1feedback prediction 1user experience evaluation 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CL2026
UXBench: Benchmarking User Experience in AI Assistants
Mengze Hong, Xia Zeng, Zeyang Lei +26
UXBench is a user‑centric benchmark that uses real interaction logs to evaluate how well AI assistants align with user preferences and generate engaging dialogue, featuring three t…
cs.CL2025
ECLM: Entity Level Language Model for Spoken Language Understanding with Chain of Intent
Shangjian Yin, Peijie Huang, Jiatian Chen +2
Large Language Models (LLMs) have demonstrated impressive capabilities in language generation and general task performance. However, their application to spoken language understand…
cs.CL2024
Self-supervised Preference Optimization: Enhance Your Language Model with Preference Degree Awareness
Jian Li, Haojing Huang, Yujia Zhang +6
Recently, there has been significant interest in replacing the reward model in Reinforcement Learning with Human Feedback (RLHF) methods for Large Language Models (LLMs), such as D…