agentic reinforcement learning 1hierarchical reward modeling 1sandbox simulation 1tool-use agents 1travel planning 1
From the 1 of 2 linked papers with an AI index.
2 papers
astro-ph.CO2026
Forecast for the detectability of patchy hydrogen reionization in WEAVE-QSO measurements of the Lyman- forest power spectrum at redshift
Ke Ma, James S. Bolton, Vid Iršič +9
We present the first detailed forecasts for the detectability of patchy hydrogen reionization in the one-dimensional Ly forest power spectrum to be measured by the WEAVE-QSO sur…
cs.AI2026
DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents
Yansong Ning, Rui Liu, Jun Wang +6
The paper introduces DeepTravel, an end‑to‑end reinforcement‑learning framework that trains autonomous travel‑planning agents to plan itineraries, invoke external tools, and self‑c…