ai assistants 1benchmark dataset 1dialogue systems 1feedback prediction 1user experience evaluation 1
From the 1 of 9 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis
Xinhao Yao, Yuanzhuo Liu, Changhao Wang +6
Deep search is becoming a core capability of modern agent systems, yet it is typically evaluated solely based on end-to-end answer accuracy. This coupled evaluation paradigm entang…
cs.AI2025
Structural Reward Model: Enhancing Interpretability, Efficiency, and Scalability in Reward Modeling
Xiaoyu Liu, Di Liang, Chang Dai +9
Reward Models (RMs) are key components for evaluating and guiding language model outputs. However, traditional scalar RMs often struggle with incorporating contextual and backgroun…