3 papers
cs.AI2026
LLM Personas as a Substitute for Field Experiments in Method Benchmarking
Enoch Hyunwook Kang
Field experiments (A/B tests) are often the most credible benchmark for methods (algorithms) in societal systems, but their cost and latency bottleneck rapid methodological progres…
cs.AI2025
TextBO: Bayesian Optimization in Language Space for Eval-Efficient Self-Improving AI
Enoch Hyunwook Kang, Hema Yoganarasimhan
Large Language Models (LLMs) have enabled self-improving AI systems that iteratively generate, evaluate, and refine their outcomes. Recent studies show that prompt-optimization-bas…
cs.LG2025
Stability and Generalization for Bellman Residuals
Enoch H. Kang, Kyoungseok Jang
Offline reinforcement learning and offline inverse reinforcement learning aim to recover near-optimal value functions or reward models from a fixed batch of logged trajectories, ye…