1 citations · 1 across the 4 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
cs.CL2025
RoleRMBench & RoleRM: Towards Reward Modeling for Profile-Based Role Play in Dialogue Systems
Hang Ding, Qiming Feng, Dongqi Liu +9
Reward modeling has become a cornerstone of aligning large language models (LLMs) with human preferences. Yet, when extended to subjective and open-ended domains such as role play,…
cs.CV2025
MARRS: Masked Autoregressive Unit-based Reaction Synthesis
Yabiao Wang, Shuo Wang, Jiangning Zhang +3
This work aims at a challenging task: human action-reaction synthesis, i.e., generating human reactions conditioned on the action sequence of another person. Currently, autoregress…