5 papers
PragReST: Self-Reinforcing Counterfactual Reasoning for Pragmatic Language Understanding
Jihyung Park, Minchao Huang, Leqi Liu +1
Natural language understanding often depends on meanings that are implied rather than explicitly stated, requiring pragmatic reasoning. Despite strong performance on math and logic…
AERIC: Anticipatory Hidden-State Monitoring for Implicit Harmful Dialogue
Jihyung Park, Saleh Afroogh, Junfeng Jiao
Current language models create two safety challenges: risk must be detected early enough to avoid exposing harmful continuation, and the harmfulness itself may be implicit rather t…
Do You Feel Comfortable? Detecting Hidden Conversational Escalation in AI Chatbots
Jihyung Park, Saleh Afroogh, David Atkinson +1
Large Language Models (LLM) are increasingly integrated into everyday interactions, serving not only as information assistants but also as emotional companions. Even in the absence…
SafeMate: A Modular RAG-Based Agent for Context-Aware Emergency Guidance
Junfeng Jiao, Jihyung Park, Yiming Xu +2
Despite the abundance of public safety documents and emergency protocols, most individuals remain ill-equipped to interpret and act on such information during crises. Traditional e…
KFinEval-Pilot: A Comprehensive Benchmark Suite for Korean Financial Language Understanding
Bokwang Hwang, Seonkyu Lim, Taewoong Kim +24
We introduce KFinEval-Pilot, a benchmark suite specifically designed to evaluate large language models (LLMs) in the Korean financial domain. Addressing the limitations of existing…