Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Adversarial Training for Process Reward Models
Gurusha Juneja, Deepak Nathani, William Yang Wang
Process Reward Models (PRMs) enhance reasoning ability of LLMs by providing step-level supervision. However, their widespread adoption is limited due to expensive manual step-level…
cs.LG2024
A Mathematical Framework and a Suite of Learning Techniques for Neural-Symbolic Systems
Charles Dickens, Connor Pryor, Changyu Gao +5
The field of Neural-Symbolic (NeSy) systems is growing rapidly. Proposed approaches show great promise in achieving symbiotic unions of neural and symbolic methods. However, a unif…