From the 1 of 18 linked papers with an AI index.
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
LaGO: Latent Action Guidance for Online Reinforcement Learning
Kuan-Yen Liu, Ren-Jyun Huang, Ti-Rong Wu
Large language models (LLMs) have shown strong potential for planning and sequential decision-making, but prior work often relies on using them as direct controllers, which require…
cs.AI2026
WallZero: Mastering the Game of WallGo with Strategic Analysis
Hsing-Yu Chen, Jérôme Arjonilla, I-Chen Wu +1
WallGo is a recently introduced strategic board game popularized by the 2025 Netflix series The Devil's Plan. Although played on a small 7 x 7 board, its combination of stone movem…
cs.AI2026
Beyond Length Scaling: Synergizing Breadth and Depth for Generative Reward Models
Qiyuan Zhang, Yufei Wang, Tianhe Wu +5
Recent advancements in Generative Reward Models (GRMs) have demonstrated that scaling the length of Chain-of-Thought (CoT) reasoning considerably enhances the reliability of evalua…