7 papers
Localized Adaptation Reveals Distinct Learning Signatures in Transformers
Rebecca Ramnauth, Brian Scassellati
Transformer adaptation is typically distributed across model depth, even when the intended change is narrow. We investigate how adaptation site shapes what a model learns, how well…
The Attentional White Bear Effect in Transformer Language Models
Rebecca Ramnauth, Brian Scassellati
Instruction-based suppression is widely used to prevent language models from generating prohibited content, yet it remains unclear whether suppression reduces internal representati…
Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains
Rebecca Ramnauth, Drazen Brscic, Brian Scassellati
Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cumulative and context-dependen…
A Robot-Assisted Approach to Small Talk Training for Adults with ASD
Rebecca Ramnauth, Dražen BrÅ¡ÄiÄ, Brian Scassellati
From dating to job interviews, making new friends or simply chatting with the cashier at checkout, engaging in small talk is a vital, everyday social skill. For adults with Autism…
Gaze Behavior During a Long-Term, In-Home, Social Robot Intervention for Children with ASD
Rebecca Ramnauth, Frederick Shic, Brian Scassellati
Atypical gaze behavior is a diagnostic hallmark of Autism Spectrum Disorder (ASD), playing a substantial role in the social and communicative challenges that individuals with ASD f…
A Grounded Observer Framework for Establishing Guardrails for Foundation Models in Socially Sensitive Domains
Rebecca Ramnauth, Dražen BrÅ¡ÄiÄ, Brian Scassellati
As foundation models increasingly permeate sensitive domains such as healthcare, finance, and mental health, ensuring their behavior meets desired outcomes and social expectations…