9 papers
Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM Collectives
Changgeon Ko, Jisu Shin, Hoyun Song +3
Large language model (LLM) agents are increasingly acting as human delegates in multi-agent environments, where a representative agent integrates diverse peer perspectives to make…
Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation
Huije Lee, Jisu Shin, Hoyun Song +2
Static benchmarks for harmful content detection face limitations in scalability and diversity, and may also be affected by contamination from web-scale pre-training corpora. To add…
Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models
Seungho Cho, Changgeon Ko, Eui Jun Hwang +3
Large language models (LLMs) are increasingly used across diverse cultural contexts, making accurate cultural understanding essential. Prior evaluations have mostly focused on outp…
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
Fitsum Gaim, Hoyun Song, Huije Lee +3
Content moderation research has recently made significant advances, but remains limited in serving the majority of the world's languages due to the lack of resources, leaving milli…
Temporal Information Retrieval via Time-Specifier Model Merging
SeungYoon Han, Taeho Hwang, Sukmin Cho +4
The rapid expansion of digital information and knowledge across structured and unstructured sources has heightened the importance of Information Retrieval (IR). While dense retriev…
Does Rationale Quality Matter? Enhancing Mental Disorder Detection via Selective Reasoning Distillation
Hoyun Song, Huije Lee, Jisu Shin +3
The detection of mental health problems from social media and the interpretation of these results have been extensively explored. Research has shown that incorporating clinical sym…