1 paper · 1 filter
Haoqin Tu, Yunhao Fang, Yizhong Wang +2
Humans continuously learn from experience, whereas conventional large language model (LLM) evaluations ignore the models' ability to improve through inference-time interaction. In…