Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Benchmarking Agentic Review Systems
Dang Nguyen, Wanqing Hao, Yanai Elazar +1
A new class of agentic review systems are emerging as a remedy to the pressure placed on peer review systems by AI-assisted research, but it is unclear how they should be evaluated…
cs.AI2026
Iterative Finetuning is Mostly Idempotent
Zephaniah Roe, Jack Sanderson, Dang Nguyen +5
If a model has some behavioral tendency, such as sycophancy or misalignment, and it is trained on its own outputs, will the tendency be amplified in the next generation of models?…