2 papers
cs.AI2025
Is PRM Necessary? Problem-Solving RL Implicitly Induces PRM Capability in LLMs
Zhangying Feng, Qianglong Chen, Ning Lu +6
The development of reasoning capabilities represents a critical frontier in large language models (LLMs) research, where reinforcement learning (RL) and process reward models (PRMs…
cs.CE2025
SimLOB: Learning Representations of Limited Order Book for Financial Market Simulation
Yuanzhe Li, Yue Wu, Muyao Zhong +2
Financial market simulation (FMS) serves as a promising tool for understanding market anomalies and the underlying trading behaviors. To ensure high-fidelity simulations, it is cru…