2 papers
stat.ML2026
An Efficient Algorithm for Thresholding Monte Carlo Tree Search
Shoma Nameki, Atsuyoshi Nakamura, Junpei Komiyama +1
We introduce the Thresholding Monte Carlo Tree Search problem, in which, given a tree and a threshold , a player must answer whether the root node value of $\math…
cs.RO2025
Sim-Anchored Learning for On-the-Fly Adaptation
Bassel El Mabsout, Shahin Roozkhosh, Siddharth Mysore +2
Fine-tuning simulation-trained RL agents with real-world data often degrades crucial behaviors due to limited or skewed data distributions. We argue that designer priorities exist…