From the 1 of 4 linked papers with an AI index.
4 papers
Efficient Test-Time Optimization for Multi-Agent Proof Autoformalization
Tian-Shuo Liu, Shiyuan Zhang, Zijie Geng +5
The paper introduces ToMap, a multi‑agent system that treats proof autoformalization as a Decomposer‑Formalizer‑Prover pipeline and concentrates test‑time optimization on improving…
A Survey on Large Language Models for Mathematical Reasoning
Peng-Yuan Wang, Tian-Shuo Liu, Chenyang Wang +8
Mathematical reasoning has long represented one of the most fundamental and challenging frontiers in artificial intelligence research. In recent years, large language models (LLMs)…
Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
Xu-Hui Liu, Tian-Shuo Liu, Shengyi Jiang +4
Combining offline and online reinforcement learning (RL) techniques is indeed crucial for achieving efficient and safe learning where data acquisition is expensive. Existing method…
: Discovering Logical Formulaic Alphas using Deep Reinforcement Learning
Feng Xu, Yan Yin, Xinyu Zhang +3
Alphas are pivotal in providing signals for quantitative trading. The industry highly values the discovery of formulaic alphas for their interpretability and ease of analysis, comp…