2 papers
cs.CL2025
Feedback Forensics: A Toolkit to Measure AI Personality
Arduin Findeis, Timo Kaufmann, Eyke Hüllermeier +1
Some traits making a "good" AI model are hard to describe upfront. For example, should responses be more polite or more casual? Such traits are sometimes summarized as model charac…
cs.LG2025
Learning from Preferences and Mixed Demonstrations in General Settings
Jason R Brown, Carl Henrik Ek, Robert D Mullins
Reinforcement learning is a general method for learning in sequential settings, but it can often be difficult to specify a good reward function when the task is complex. In these c…