2 papers
cs.LG2026
PokeRL: Reinforcement Learning for Pokemon Red
Dheeraj Mudireddy, Sai Patibandla
Pokemon Red is a long-horizon JRPG with sparse rewards, partial observability, and quirky control mechanics that make it a challenging benchmark for reinforcement learning. While r…
cs.MA2025
ThinkTank: A Framework for Generalizing Domain-Specific AI Agent Systems into Universal Collaborative Intelligence Platforms
Praneet Sai Madhu Surabhi, Dheeraj Reddy Mudireddy, Jian Tao
This paper presents ThinkTank, a comprehensive and scalable framework designed to transform specialized AI agent systems into versatile collaborative intelligence platforms capable…