1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CL2026
Disparities In Negation Understanding Across Languages In Vision-Language Models
Charikleia Moraitaki, Sarah Pan, Skyler Pulling +3
Vision-language models (VLMs) exhibit affirmation bias: a systematic tendency to select positive captions ("X is present") even when the correct description contains negation ("no…
cs.CL2025
Tiny Reward Models
Sarah Pan
Large decoder-based language models have become the dominant architecture for reward modeling in reinforcement learning from human feedback (RLHF). However, as reward models are in…
cs.CL2023★ 1 cited
Let's Reinforce Step by Step
Sarah Pan, Vladislav Lialin, Sherin Muckatira +1
While recent advances have boosted LM proficiency in linguistic benchmarks, LMs consistently struggle to reason correctly on complex tasks like mathematics. We turn to Reinforcemen…