10 citations · 11 across the 4 of their papers we have counts for
4 papers
Do Chess Explanations Reflect Model Decisions? Behavioral and Token-Level Tests of LLM Reasoning Faithfulness
Angelina Parfenova
Large language models can produce fluent explanations for chess moves, but plausible language does not necessarily reflect the reasoning behind a decision. We study this question i…
From Quotes to Concepts: Axial Coding of Political Debates with Ensemble LMs
Angelina Parfenova, David Graus, Juergen Pfeffer
Axial coding is a commonly used qualitative analysis method that enhances document understanding by organizing sentence-level open codes into broader categories. In this paper, we…
Emergent Convergence in Multi-Agent LLM Annotation
Angelina Parfenova, Alexander Denzler, Juergen Pfeffer
Large language models (LLMs) are increasingly deployed in collaborative settings, yet little is known about how they coordinate when treated as black-box agents. We simulate 7500 m…
Text Annotation via Inductive Coding: Comparing Human Experts to LLMs in Qualitative Data Analysis
Angelina Parfenova, Andreas Marfurt, Alexander Denzler +1
This paper investigates the automation of qualitative data analysis, focusing on inductive coding using large language models (LLMs). Unlike traditional approaches that rely on ded…