Showing 2024Show all
2 papers · 1 filter
cs.CL2024
MetaRuleGPT: Recursive Numerical Reasoning of Language Models Trained with Simple Rules
Kejie Chen, Lin Wang, Qinghai Zhang +1
Recent studies have highlighted the limitations of large language models in mathematical reasoning, particularly their inability to capture the underlying logic. Inspired by meta-l…
cs.LG2024
Goal Exploration via Adaptive Skill Distribution for Goal-Conditioned Reinforcement Learning
Lisheng Wu, Ke Chen
Exploration efficiency poses a significant challenge in goal-conditioned reinforcement learning (GCRL) tasks, particularly those with long horizons and sparse rewards. A primary li…