2 papers
cs.LG2026
Not only where, But when: Temporal Scheduling for RLVR
Jinghao Zhang, Ruilin Li, Feng Zhao +1
Reinforcement learning with verifiable rewards (RLVR) has become a core technique for post-training of Large Language Models (LLMs). While policy optimization is driven by all samp…
cs.LG2025
Beyond Tree Models: A Hybrid Model of KAN and gMLP for Large-Scale Financial Tabular Data
Mingming Zhang, Jiahao Hu, Pengfei Shi +8
Tabular data plays a critical role in real-world financial scenarios. Traditionally, tree models have dominated in handling tabular data. However, financial datasets in the industr…