Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Causal Intervention-Based Memory Selection for Long-Horizon LLM Agents
Saksham Sahai Srivastava
Long-horizon LLM agents rely on persistent memory to support interactions across sessions, yet existing memory systems often retrieve context using semantic similarity or broad his…
cs.AI2025
A Technical Survey of Reinforcement Learning Techniques for Large Language Models
Saksham Sahai Srivastava, Vaneet Aggarwal
This survey offers a comprehensive foundation on the integration of RL with language models, highlighting prominent algorithms such as Proximal Policy Optimization (PPO), Q-Learnin…