Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
ProCUA-SFT Technical Report
Jaehun Jung, Ximing Lu, Brandon Cui +11
Training computer-use agents (CUAs) -- models that interact with graphical desktops through screenshots and keyboard/mouse actions -- requires large-scale, diverse trajectory data…
cs.LG2026
Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking in RLVR
Muhammad Khalifa, Zohaib Khan, Omer Tafveez +2
Reward hacking is a form of misalignment in which models overoptimize proxy rewards without genuinely solving the underlying task. Precisely measuring reward hacking occurrence rem…