3 papers
cs.AI2026
Open Problems in Constitutional Preference Reconstruction
Eleanor Clifford, Michael Amir, Arduin Findeis +2
Pairwise preference data is widely used for training and evaluating language models (e.g., RLHF), but each datapoint records a \emph{choice}, not the rationale behind it. Methods s…
cs.CL2025
Feedback Forensics: A Toolkit to Measure AI Personality
Arduin Findeis, Timo Kaufmann, Eyke Hüllermeier +1
Some traits making a "good" AI model are hard to describe upfront. For example, should responses be more polite or more casual? Such traits are sometimes summarized as model charac…
cs.CL2025
Can External Validation Tools Improve Annotation Quality for LLM-as-a-Judge?
Arduin Findeis, Floris Weers, Guoli Yin +3
Pairwise preferences over model responses are widely collected to evaluate and provide feedback to large language models (LLMs). Given two alternative model responses to the same i…