4 papers
Perfecting Human-AI Interaction at Clinical Scale. Turning Production Signals into Safer, More Human Conversations
Subhabrata Mukherjee, Markel Sanz Ausin, Kriti Aggarwal +24
Healthcare conversational AI agents shouldn't be optimized only for clean benchmark accuracy in production-first regime; they must be optimized for the lived reality of patient con…
NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment
Gerald Shen, Zhilin Wang, Olivier Delalleau +10
Aligning Large Language Models (LLMs) with human values and preferences is essential for making them helpful and safe. However, building efficient tools to perform alignment can be…
Polaris: A Safety-focused LLM Constellation Architecture for Healthcare
Subhabrata Mukherjee, Paul Gamble, Markel Sanz Ausin +23
We develop Polaris, the first safety-focused LLM constellation for real-time patient-AI healthcare conversations. Unlike prior LLM works in healthcare focusing on tasks like questi…
InferNet for Delayed Reinforcement Tasks: Addressing the Temporal Credit Assignment Problem
Markel Sanz Ausin, Hamoon Azizsoltani, Song Ju +2
The temporal Credit Assignment Problem (CAP) is a well-known and challenging task in AI. While Reinforcement Learning (RL), especially Deep RL, works well when immediate rewards ar…