4 papers
Differentiable Evolutionary Reinforcement Learning
Sitao Cheng, Tianle Li, Xuhan Huang +2
Crafting effective reward signals remains a central challenge in Reinforcement Learning (RL), especially for complex reasoning tasks. Existing automated reward optimization methods…
CALM Before the STORM: Unlocking Native Reasoning for Optimization Modeling
Zhengyang Tang, Zihan Ye, Chenyu Huang +9
Large Reasoning Models (LRMs) have demonstrated strong capabilities in complex multi-step reasoning, opening new opportunities for automating optimization modeling. However, existi…
Federated Linear Dueling Bandits
Xuhan Huang, Yan Hu, Zhiyan Li +3
Contextual linear dueling bandits have recently garnered significant attention due to their widespread applications in important domains such as recommender systems and large langu…
LLMs for Mathematical Modeling: Towards Bridging the Gap between Natural and Mathematical Languages
Xuhan Huang, Qingning Shen, Yan Hu +2
Large Language Models (LLMs) have demonstrated strong performance across various natural language processing tasks, yet their proficiency in mathematical reasoning remains a key ch…