Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Bayesian Symbolic Regression with Entropic Reinforcement Learning
Oussama Boussif, Mohammed Mahfoud, Younesse Kaddar +6
Symbolic regression is the problem of finding an algebraic expression describing a stochastic dependence of a target variable on a set of inputs. Unlike forms of regression that fi…
cs.LG2026
A Measure-Theoretic Analysis of Reasoning: Structural Generalization and Approximation Limits
Yuyang Zhang, Yifu Zhang, Xuehai Zhou +1
While empirical scaling laws for LLM reasoning are well-documented, the theoretical mechanisms governing out-of-distribution (OOD) generalization remain elusive. We formalize reaso…
cs.LG2025
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
Sanxing Chen, Xiaoyin Chen, Yukun Huang +2
While Large Language Models (LLMs) hold promise to become autonomous agents, they often explore suboptimally in sequential decision-making. Recent work has sought to enhance this c…