collaborators

12 papers

cs.AI2026

EduZone: A Framework for Evaluating LLM Safety for K-12 Students and Teachers

Junyeong Park, Jieun Han, Haneul Yoo +3

Large language models (LLMs) are increasingly used across diverse tasks in K-12 education, yet existing safety evaluations rarely examine how harmful or inappropriate content appea…

cs.CL2026

IterCOMP: Reasoning-aware Adaptive Prompt Compression for Multi-hop Question Answering

JungMin Yun, YoungBin Kim

Multi-hop question answering requires complex reasoning across multiple evidence segments, which often overwhelms retrieval-augmented generation systems with lengthy and noisy cont…

cs.CL2026

On the Effect of Uncertainty on Layer-wise Inference Dynamics

Sunwoo Kim, Haneul Yoo, Alice Oh

Understanding how large language models (LLMs) internally represent and process their predictions is central to detecting uncertainty and preventing hallucinations. While several s…

cs.CL2026

SemEval-2026 Task 7: Everyday Knowledge Across Diverse Languages and Cultures

Nedjma Ousidhoum, Junho Myung, Carla Perez-Almendros +27

We present our shared task on evaluating the adaptability of LLMs and NLP systems across multiple languages and cultures. The task data consist of an extended version of our manual…

cs.CL2026

FINEST: Improving LLM Responses to Sensitive Topics Through Fine-Grained Evaluation

Juhyun Oh, Nayeon Lee, Chani Jung +5

Large Language Models (LLMs) often generate overly cautious and vague responses on sensitive topics, sacrificing helpfulness for safety. Existing evaluation frameworks lack systema…

cs.HC2026

When Scaffolding Breaks: Investigating Student Interaction with LLM-Based Writing Support in Real-Time K-12 EFL Classrooms

Junho Myung, Hyunseung Lim, Hana Oh +6

Large language models (LLMs) are promising tools for scaffolding students' English writing skills, but their effectiveness in real-time K-12 classrooms remains underexplored. Addre…