5 papers
Commonsense Knowledge with Negation: A Resource to Enhance Negation Understanding
Zijie Wang, MohammadHossein Rezaei, Farzana Rashid +1
Negation is a common and important semantic feature in natural language, yet Large Language Models (LLMs) struggle when negation is involved in natural language understanding tasks…
LLM Reasoning with Process Rewards for Outcome-Guided Steps
Mohammad Rezaei, Jens Lehmann, Sahar Vahdati
Mathematical reasoning in large language models has improved substantially with reinforcement learning using verifiable rewards, where final answers can be checked automatically an…
Online Rubrics Elicitation from Pairwise Comparisons
MohammadHossein Rezaei, Robert Vacareanu, Zihao Wang +4
Rubrics provide a flexible way to train LLMs on open-ended long-form answers where verifiable rewards are not applicable and human preferences provide coarse signals. Prior work sh…
Making Language Models Robust Against Negation
MohammadHossein Rezaei, Eduardo Blanco
Negation has been a long-standing challenge for language models. Previous studies have shown that they struggle with negation in many natural language understanding tasks. In this…
EgoNormia: Benchmarking Physical Social Norm Understanding
MohammadHossein Rezaei, Yicheng Fu, Phil Cuvin +4
Human activity is moderated by norms; however, supervision for normative reasoning is sparse, particularly where norms are physically- or socially-grounded. We thus present EGONORM…