From the 1 of 9 linked papers with an AI index.
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
RetroAgent: Harnessing LLMs to Search Over Structured Memory for Agentic Retrosynthesis Planning
Yanqiao Zhu, Jingru Gan, Xiaoqi Sun +8
RetroAgent is an LLM-based agent that uses structured memory to integrate symbolic search and neural reasoning for multi-step retrosynthesis planning, enabling informed decisions a…
cs.AI2026
Can Current Agents Close the Discovery-to-Application Gap? A Case Study in Minecraft
Zhou Ziheng, Huacong Tang, Jinyuan Zhang +9
Discovering causal regularities and applying them to build functional systems--the discovery-to-application loop--is a hallmark of general intelligence, yet evaluating this capacit…
cs.AI2026
TPO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning
Haixin Wang, Hejie Cui, Chenwei Zhang +7
Recent progress in multi-turn reinforcement learning (RL) has significantly improved reasoning LLMs' performances on complex interactive tasks. Despite advances in stabilization te…