3 papers
cs.CL2025
Judging with Confidence: Calibrating Autoraters to Preference Distributions
Zhuohang Li, Xiaowei Li, Chengyu Huang +11
The alignment of large language models (LLMs) with human values increasingly relies on using other LLMs as automated judges, or ``autoraters''. However, their reliability is limite…
cs.CL2025
Handling Students Dropouts in an LLM-driven Interactive Online Course Using Language Models
Yuanchun Wang, Yiyang Fu, Jifan Yu +7
Interactive online learning environments, represented by Massive AI-empowered Courses (MAIC), leverage LLM-driven multi-agent systems to transform passive MOOCs into dynamic, text-…
cs.CV2025
The Tenth NTIRE 2025 Image Denoising Challenge Report
Lei Sun, Hang Guo, Bin Ren +91
This paper presents an overview of the NTIRE 2025 Image Denoising Challenge (Ï = 50), highlighting the proposed methodologies and corresponding results. The primary objective is t…