7 papers · 1 filter
Evidence-Consistent Generative Detection under Scenario-Level Distribution Shift
San Kim, JinYeong Bak
Conventional in-distribution evaluation can overestimate robustness when training and test data share recurring task-specific patterns or surface cues. This risk is especially rele…
MindTailor: Personalized Emotional Support via Post History-Grounded Case Formulation and Collaborative Refinement
Suhyun Han, Kyunghyun Cho, JinYeong Bak
As mental health concerns continue to rise globally, social media has emerged as a vital space where individuals seek emotional support. While prior work on personalized emotional…
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
Jaehyeok Lee, Xiaoyuan Yi, Jing Yao +4
As LLMs are globally deployed, aligning their cultural value orientations is critical for safety and user engagement. However, existing benchmarks face the Construct-Composition-Co…
Camellia: Benchmarking Cultural Biases in LLMs for Asian Languages
Tarek Naous, Anagha Savit, Carlos Rafael Catalan +17
As Large Language Models (LLMs) develop stronger multilingual capabilities, their sensitivity to culturally diverse entities becomes increasingly important. Prior work by Naous et…
KpopMT: Translation Dataset with Terminology for Kpop Fandom
JiWoo Kim, Yunsu Kim, JinYeong Bak
While machines learn from existing corpora, humans have the unique capability to establish and accept new language systems. This makes human form unique language systems within soc…
MentalAgora: A Gateway to Advanced Personalized Care in Mental Health through Multi-Agent Debating and Attribute Control
Yeonji Lee, Sangjun Park, Kyunghyun Cho +1
As mental health issues globally escalate, there is a tremendous need for advanced digital support systems. We introduce MentalAgora, a novel framework employing large language mod…