13 citations · 17 across the 4 of their papers we have counts for
4 papers
Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Zhiyuan Zeng, Qinyuan Cheng, Zhangyue Yin +6
OpenAI o1 represents a significant milestone in Artificial Inteiligence, which achieves expert-level performances on many challanging tasks that require strong reasoning ability.Op…
EvoWiki: Evaluating LLMs on Evolving Knowledge
Wei Tang, Yixin Cao, Yang Deng +8
Knowledge utilization is a critical aspect of LLMs, and understanding how they adapt to evolving knowledge is essential for their effective deployment. However, existing benchmarks…
Constraints on large-scale polarization in northern hemisphere
Dongdong Zhang, Bo Wang, Jia-Rui Li +2
Present cosmic microwave background (CMB) observations have significantly advanced our understanding of the universe's origin, especially with primordial gravitational waves (PGWs)…
Magnetic Field-Based Reward Shaping for Goal-Conditioned Reinforcement Learning
Hongyu Ding, Yuanze Tang, Qing Wu +3
Goal-conditioned reinforcement learning (RL) is an interesting extension of the traditional RL framework, where the dynamic environment and reward sparsity can cause conventional l…