Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Reward Under Attack: Analyzing the Robustness and Hackability of Process Reward Models
Rishabh Tiwari, Aditya Tomar, Udbhav Bamba +5
Process Reward Models (PRMs) are rapidly becoming the backbone of LLM reasoning pipelines, yet we demonstrate that state-of-the-art PRMs are systematically exploitable under advers…
cs.LG2024
Control-oriented Clustering of Visual Latent Representation
Han Qi, Haocheng Yin, Heng Yang
We initiate a study of the geometry of the visual representation space -- the information channel from the vision encoder to the action decoder -- in an image-based control pipelin…