5 papers · 1 filter
Lost in Historical Time? A Polish History Matura Benchmark for Large Language Models
Adrian Trzoss, Kacper Dudzic, Wiktor Werner +1
Language models are widely used by students as knowledge sources, yet benchmarks rarely assess their interpretative historical reasoning. We evaluate eight leading LLMs on the Poli…
The Two-Process Theory of Machine Self-Report
Hubert Plisiecki, Filip Chmielewski, Kacper Dudzic +3
Language models are increasingly asked to self-report, informing safety evaluations, public understanding, and model-welfare debates. Yet their reports are elicited with human ques…
Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives
Karolina Drożdż, Kacper Dudzic, Anna Sterna +1
Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. This depth over breadth case study direct…
The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences
Hubert Plisiecki, Sabina Siudaj, Kacper Dudzic +4
We administer 45 validated psychometric questionnaires to 50 large language models (LLMs) to identify the dimensions along which LLMs differ psychometrically. Using Supervised Sema…
Computational Phenomenology of Temporal Experience in Autism: Quantifying the Emotional and Narrative Characteristics of Lived Unpredictability
Kacper Dudzic, Karolina Drożdż, Maciej WodziÅski +2
Disturbances in temporality, such as desynchronization with the social environment and its unpredictability, are considered core features of autism with a deep impact on relationsh…