2 papers
cs.CL2026
Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis
Hyeji Choi, Yongtaek Lim, Minwoo Kim
Multilingual safety evaluation of large language models (LLMs) has predominantly relied on direct translation (DT) of English benchmarks into target languages - an approach that co…
cs.AI2026
Reliable to Expressive: A Curriculum for Rubric-Following Safety Judges
Yongtaek Lim, Hyeji Choi, Minwoo Kim
Safety judges are increasingly deployed to evaluate model outputs against evolving criteria, yet recent meta-evaluation work shows they remain brittle under prompt and rubric varia…