adaptive learning 1agent architecture 1autonomous systems 1foundation models 1self-improving agents 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.AI2026
Self-Improvements in Modern Agentic Systems: A Survey
Zhe Ren, Yimeng Chen, Dandan Guo +9
The paper surveys modern self-improving autonomous agents, presenting a system-level framework that combines foundation models with prompts, memory, tools, and control logic, and c…
cs.LG2026
Mitigating Reward Hacking in RLHF via Bayesian Non-negative Reward Modeling
Zhibin Duan, Guowei Rong, Zhuo Li +3
Reward models learned from human preferences are central to aligning large language models (LLMs) via reinforcement learning from human feedback, yet they are often vulnerable to r…
cs.LG2025
Merging Smarter, Generalizing Better: Enhancing Model Merging on OOD Data
Bingjie Zhang, Hongkang Li, Changlong Shi +5
Multi-task learning (MTL) concurrently trains a model on diverse task datasets to exploit common features, thereby improving overall performance across the tasks. Recent studies ha…