6 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CL2025
Reward Models are Metrics in a Trench Coat
Sebastian Gehrmann
The emergence of reinforcement learning in post-training of large language models has sparked significant interest in reward models. Reward models assess the quality of sampled mod…
cs.CL2025★ 6 cited
Understanding and Mitigating Risks of Generative AI in Financial Services
Sebastian Gehrmann, Claire Huang, Xian Teng +9
To responsibly develop Generative AI (GenAI) products, it is critical to define the scope of acceptable inputs and outputs. What constitutes a "safe" response is an actively debate…