2 papers
cs.CL2025
R.U.Psycho? Robust Unified Psychometric Testing of Language Models
Julian Schelb, Orr Borin, David Garcia +1
Generative language models are increasingly being subjected to psychometric questionnaires intended for human testing, in efforts to establish their traits, as benchmarks for align…
cs.CY2025
Only a Little to the Left: A Theory-grounded Measure of Political Bias in Large Language Models
Mats Faulborn, Indira Sen, Max Pellert +2
Prompt-based language models like GPT4 and LLaMa have been used for a wide variety of use cases such as simulating agents, searching for information, or for content analysis. For a…