2 papers
cs.LG2025
Optimizing Prompt Sequences using Monte Carlo Tree Search for LLM-Based Optimization
Fei Xu Yu, Gina Adam, Nathaniel D. Bastian +1
Large language models (LLMs) have demonstrated remarkable capabilities in code generation and structured reasoning; however, their performance often degrades on complex tasks that…
cs.LG2024
RGMDT: Return-Gap-Minimizing Decision Tree Extraction in Non-Euclidean Metric Space
Jingdi Chen, Hanhan Zhou, Yongsheng Mei +4
Deep Reinforcement Learning (DRL) algorithms have achieved great success in solving many challenging tasks while their black-box nature hinders interpretability and real-world appl…