3 papers
cs.CL2025
Incremental Summarization for Customer Support via Progressive Note-Taking and Agent Feedback
Yisha Wu, Cen Mia Zhao, Yuanpei Cao +4
We introduce an incremental summarization system for customer support agents that intelligently determines when to generate concise bullet notes during conversations, reducing agen…
cs.LG2025
ObjectRL: An Object-Oriented Reinforcement Learning Codebase
Gulcin Baykal, Abdullah Akgül, Manuel Haussmann +4
ObjectRL is an open-source Python codebase for deep reinforcement learning (RL), designed for research-oriented prototyping with minimal programming effort. Unlike existing codebas…
cs.LG2025
Deep Actor-Critics with Tight Risk Certificates
Bahareh Tasdighi, Manuel Haussmann, Yi-Shan Wu +2
Deep actor-critic algorithms have reached a level where they influence everyday life. They are a driving force behind continual improvement of large language models through user fe…