1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2026
Reinforcement Learning Towards Broadly and Persistently Beneficial Models
Akshay V. Jagadeesh, Rahul K. Arora, Khaled Saab +5
As AI systems are deployed across increasingly diverse and high-stakes settings, model alignment must generalize beyond the tasks and domains seen during training. This is especial…
cs.CL2026★ 1 cited
HealthBench Professional: Evaluating Large Language Models on Real Clinician Chats
Rebecca Soskin Hicks, Mikhail Trofimov, Dominick Lim +13
Millions of clinicians use ChatGPT to support clinical care, but evaluations of the most common use cases in model-clinician conversations are limited. We introduce HealthBench Pro…