3 papers
cs.LG2026
A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions
Sanghwa Kim, Junghyun Lee, Se-Young Yun
We consider the problem of heteroskedastic generalized linear bandits (GLBs) with adversarial corruptions, which subsumes heteroskedastic linear bandits and logistic/Poisson bandit…
stat.ML2025
On the Optimality of Tracking Fisher Information in Adaptive Testing with Stochastic Binary Responses
Sanghwa Kim, Dohyun Ahn, Seungki Min
We study the problem of estimating a continuous ability parameter from sequential binary responses by actively asking questions with varying difficulties, a setting that arises nat…
cs.CL2025
Pre-Storage Reasoning for Episodic Memory: Shifting Inference Burden to Memory for Personalized Dialogue
Sangyeop Kim, Yohan Lee, Sanghwa Kim +2
Effective long-term memory in conversational AI requires synthesizing information across multiple sessions. However, current systems place excessive reasoning burden on response ge…