3 papers
cs.AI2026
Multi-Turn Reinforcement Learning for Tool-Calling Agents with Iterative Reward Calibration
Wachiravit Modecrua, Krittanon Kaewtawee, Krittin Pachtrachai +1
Training tool-calling agents with reinforcement learning on multi-turn tasks remains challenging due to sparse outcome rewards and difficult credit assignment across conversation t…
cs.AI2025
ROAD: Reflective Optimization via Automated Debugging for Zero-Shot Agent Alignment
Natchaya Temyingyong, Daman Jain, Neeraj Kumarsahu +6
Automatic Prompt Optimization (APO) has emerged as a critical technique for enhancing Large Language Model (LLM) performance, yet current state-of-the-art methods typically rely on…
cs.AI2025
Cloning a Conversational Voice AI Agent from Call\,Recording Datasets for Telesales
Krittanon Kaewtawee, Wachiravit Modecrua, Krittin Pachtrachai +1
Recent advances in language and speech modelling have made it possible to build autonomous voice assistants that understand and generate human dialogue in real time. These systems…