works on

From the 1 of 5 linked papers with an AI index.

collaborators

5 papers

cs.CL2026

K-EXAONE 2.0 Technical Report

Eunbi Choi, Kibong Choi, Sehyun Chun +74

This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward global frontier-scale foundatio…

cs.LG2026

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache

Donghyun Son, Euntae Choi, Sungjoo Yoo

The paper proposes NSNQuant, a calibration‑free method that uses a double normalization and Hadamard transform to compress the key‑value cache of large language models with low‑bit…

cs.CL2026

EntropyCache: Decoded Token Entropy Guided KV Caching for Diffusion Language Models

Minsoo Cheong, Donghyun Son, Woosang Lim +1

Diffusion-based large language models (dLLMs) rely on bidirectional attention, which prevents lossless KV caching and requires a full forward pass at every denoising step. Existing…

cs.DC2026

Nalar: An agent serving framework

Marco Laju, Donghyun Son, Saurabh Agarwal +4

LLM-driven agentic applications increasingly automate complex, multi-step tasks, but serving them efficiently remains challenging due to heterogeneous components, dynamic and model…

cs.CL2024

In-Context Learning with Noisy Labels

Junyong Kang, Donghyun Son, Hwanjun Song +1

In-context learning refers to the emerging ability of large language models (LLMs) to perform a target task without additional training, utilizing demonstrations of the task. Recen…