Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Resource-constrained Amazons chess decision framework integrating large language models and graph attention
Tianhao Qian, Zhuoxuan Li, Jinde Cao +2
Artificial intelligence has advanced significantly through the development of intelligent game-playing systems, providing rigorous testbeds for decision-making, strategic planning,…
cs.AI2025
MAPO: Mixed Advantage Policy Optimization
Wenke Huang, Quan Zhang, Yiyang Fang +11
Recent advances in reinforcement learning for foundation models, such as Group Relative Policy Optimization (GRPO), have significantly improved the performance of foundation models…