2 papers
cs.CL2025
Reward Models are Metrics in a Trench Coat
Sebastian Gehrmann
The emergence of reinforcement learning in post-training of large language models has sparked significant interest in reward models. Reward models assess the quality of sampled mod…
cs.CL2025
Understanding and Mitigating Risks of Generative AI in Financial Services
Sebastian Gehrmann, Claire Huang, Xian Teng +9
To responsibly develop Generative AI (GenAI) products, it is critical to define the scope of acceptable inputs and outputs. What constitutes a "safe" response is an actively debate…