6 citations
- Beijing Academy of Artificial IntelligenceCN11 papers
- Shanghai Jiao Tong UniversityCN7 papers
- Tsinghua UniversityCN4 papers
- Chinese University of Hong KongHK3 papers
- Fudan UniversityCN3 papers
- ShangHai JiAi Genetics & IVF InstituteCN3 papers
- Hong Kong Polytechnic UniversityHK2 papers
- ShanghaiTech UniversityCN2 papers
- University of Science and Technology of ChinaCN2 papers
- Zhejiang UniversityCN2 papers
- Alibaba Group (China)CN1 paper
- Arizona State UniversityUS1 paper
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
LABO: LLM-Accelerated Bayesian Optimization through Broad Exploration and Selective Experimentation
Zhuo Chen, Xinzhe Yuan, Jianshu Zhang +8
The high cost and data scarcity in scientific exploration have motivated the use of large language models (LLMs) as knowledge-driven components in Bayesian optimization (BO). Howev…
cs.LG2026
The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMs
Zibo Zhao, Yuanting Zha, Haipeng Zhang +1
Self-reflection capabilities emerge in Large Language Models after RL post-training, with multi-turn RL achieving substantial gains over SFT counterparts. Yet the mechanism of how…