4 papers
Future Policy Approximation for Offline Reinforcement Learning in LLM Reasoning
Minjae Oh, Yunho Choi, Dongmin Choi +1
Reinforcement learning (RL) has emerged as a key driver of post-training for complex reasoning in large language models (LLMs), yet online RL introduces substantial instability and…
Human Psychometric Questionnaires Mischaracterize LLM Behavior
Woojung Song, Dongmin Choi, Yoonah Park +3
We examine whether human psychometric questionnaires can serve as reliable tools for characterizing and predicting LLM behavior in everyday user interactions. We analyze eight open…
Value Portrait: Assessing Language Models' Values through Psychometrically and Ecologically Valid Items
Jongwook Han, Dongmin Choi, Woojung Song +2
The importance of benchmarks for assessing the values of language models has been pronounced due to the growing need of more authentic, human-aligned responses. However, existing b…
PVP: An Image Dataset for Personalized Visual Persuasion with Persuasion Strategies, Viewer Characteristics, and Persuasiveness Ratings
Junseo Kim, Jongwook Han, Dongmin Choi +3
Visual persuasion, which uses visual elements to influence cognition and behaviors, is crucial in fields such as advertising and political communication. With recent advancements i…