activity
20242026
collaborators

11 papers

cs.AI2026

PersonaTrail: Benchmarking Personalized Web Agents through Browsing Trails

Seungbin Yang, Chaewoon Ki, Dohyun Lee +2

Recent advances in large language models have enabled web agents to autonomously execute complex tasks. In practice, users frequently provide underspecified instructions, requiring…

cs.CL2026

LiveWeb-IE: A Benchmark For Online Web Information Extraction

Seungbin Yang, Jihwan Kim, Jaemin Choi +4

Web information extraction (WIE) is the task of automatically extracting data from web pages, offering high utility for various applications. The evaluation of WIE systems has trad…

cs.CL2026

Not the Example, but the Process: How Self-Generated Examples Enhance LLM Reasoning

Daehoon Gwak, Minseo Jung, Junwoo Park +4

Recent studies have shown that Large Language Models (LLMs) can improve their reasoning performance through self-generated few-shot examples, achieving results comparable to manual…

cs.CL2025

Reward-Weighted Sampling: Enhancing Non-Autoregressive Characteristics in Masked Diffusion LLMs

Daehoon Gwak, Minseo Jung, Junwoo Park +4

Masked diffusion models (MDMs) offer a promising non-autoregressive alternative for large language modeling. Standard decoding methods for MDMs, such as confidence-based sampling,…

cs.CL2025

Evaluating Automatic Speech Recognition Systems for Korean Meteorological Experts

ChaeHun Park, Hojun Cho, Jaegul Choo

This paper explores integrating Automatic Speech Recognition (ASR) into natural language query systems to improve weather forecasting efficiency for Korean meteorologists. We addre…

eess.IV2025

AMRG: Extend Vision Language Models for Automatic Mammography Report Generation

Nak-Jun Sung, Donghyun Lee, Bo Hwa Choi +1

Mammography report generation is a critical yet underexplored task in medical AI, characterized by challenges such as multiview image reasoning, high-resolution visual cues, and un…