1 citations · 1 across the 5 of their papers we have counts for
6 papers
You Are What You Read: Misalignment via In-Context Persona Induction
Kyuhee Kim, Benjamin Berczi, Cozmin Ududec
Broad misalignment has been produced by finetuning on narrow data, harmful or benign, and in context only by demonstrations of the undesirable behaviour itself. We show that benign…
ChainMark: Model-Free LLM Watermarking with Closed-Form Calibration
Chengheng Li-Chen, Kyuhee Kim
Regulatory regimes such as the EU AI Act mandate machine-readable marking of synthetic text, but existing watermark detectors rely on the generating LM and on heuristic thresholds…
Do LLMs Game Formalization? Evaluating Faithfulness in Logical Reasoning
Kyuhee Kim, Auguste Poiroux, Antoine Bosselut
Formal verification guarantees proof validity but not formalization faithfulness. For natural-language logical reasoning, where models construct axiom systems from scratch without…
Nunchi-Bench: Benchmarking Language Models on Cultural Reasoning with a Focus on Korean Superstition
Kyuhee Kim, Sangah Lee
As large language models (LLMs) become key advisors in various domains, their cultural sensitivity and reasoning skills are crucial in multicultural environments. We introduce Nunc…
KoCoNovel: Annotated Dataset of Character Coreference in Korean Novels
Kyuhee Kim, Surin Lee, Sangah Lee
In this paper, we present KoCoNovel, a novel character coreference dataset derived from Korean literary texts, complete with detailed annotation guidelines. Comprising 178K tokens…
K-Act2Emo: Korean Commonsense Knowledge Graph for Indirect Emotional Expression
Kyuhee Kim, Surin Lee, Sangah Lee
In many literary texts, emotions are indirectly conveyed through descriptions of actions, facial expressions, and appearances, necessitating emotion inference for narrative underst…