2 papers
cs.CL2026
Best-of- TTS Evaluation is Confounded by ASR Family Alignment
Taehyung Yu, Seongjae Kang
Best-of- (BoN) inference improves content consistency in zero-shot text-to-speech by selecting among multiple candidates with an automatic speech recognition (ASR) verifier. We…
cs.AI2026
PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents
Seongjae Kang, Taehyung Yu, Sung Ju Hwang
LLM agents handle user requests on behalf of organizations through tool calls and must follow the company policies stated in their system prompts. Prior work approaches this as a s…