6 papers · 1 filter
Lost in Historical Time? A Polish History Matura Benchmark for Large Language Models
Adrian Trzoss, Kacper Dudzic, Wiktor Werner +1
Language models are widely used by students as knowledge sources, yet benchmarks rarely assess their interpretative historical reasoning. We evaluate eight leading LLMs on the Poli…
The Two-Process Theory of Machine Self-Report
Hubert Plisiecki, Filip Chmielewski, Kacper Dudzic +3
Language models are increasingly asked to self-report, informing safety evaluations, public understanding, and model-welfare debates. Yet their reports are elicited with human ques…
The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences
Hubert Plisiecki, Sabina Siudaj, Kacper Dudzic +4
We administer 45 validated psychometric questionnaires to 50 large language models (LLMs) to identify the dimensions along which LLMs differ psychometrically. Using Supervised Sema…
Computational Phenomenology of Temporal Experience in Autism: Quantifying the Emotional and Narrative Characteristics of Lived Unpredictability
Kacper Dudzic, Karolina Drożdż, Maciej Wodziński +2
Disturbances in temporality, such as desynchronization with the social environment and its unpredictability, are considered core features of autism with a deep impact on relationsh…
Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality Disorder Diagnosis through First-Person Narratives
Karolina Drożdż, Kacper Dudzic, Anna Sterna +1
Growing reliance on LLMs for psychiatric self-assessment raises questions about their ability to interpret qualitative patient narratives. This depth over breadth case study direct…
Two Approaches to Diachronic Normalization of Polish Texts
Kacper Dudzic, Filip Graliński, Krzysztof Jassem +2
This paper discusses two approaches to the diachronic normalization of Polish texts: a rule-based solution that relies on a set of handcrafted patterns, and a neural normalization…