4 papers
RELIC: Revealed Principles for Learning Interpretable Composable Skills in Multi-Agent Planning
Nguyen Viet Tuan Kiet, Bui Dinh Pham, Duong Quoc Chinh +3
Multi-agent planning becomes substantially harder when agents must improve specialized decision-making skills while keeping their executable implementations private. This setting a…
Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling
Dao Tran, Duc Anh Le, Ngoc Luu +3
Test-time scaling improves language model reasoning by spending additional compute to explore multiple solution trajectories. The key challenge is to maximize accuracy while minimi…
Back to the Beginning of Heuristic Design: Bridging Code and Knowledge with LLMs
Nguyen Viet Tuan Kiet, Bui Dinh Pham, Dao Van Tung +2
Large language models (LLMs) have recently advanced automatic heuristic design (AHD) for combinatorial optimization (CO), where candidate heuristics are iteratively proposed, evalu…
MMP-A*: Multimodal Perception Enhanced Incremental Heuristic Search on Path Planning
Minh Hieu Ha, Khanh Ly Ta, Hung Phan +4
Autonomous path planning requires a synergy between global reasoning and geometric precision, especially in complex or cluttered environments. While classical A* is valued for its…