2 papers
cs.CL2026
TLRD: Teaching LLMs to Reason over Tabular Data with Tri-Level Rationale Distillation
Tianyuan Liang, Xuwei Tan, Lei Shi +6
Tabular data is a primary medium for storing real-world information, driving many industrial applications of machine learning. Traditional predictors achieve strong predictive perf…
cs.AI2025
SHARP: Synthesizing High-quality Aligned Reasoning Problems for Large Reasoning Models Reinforcement Learning
Xiong Jun Wu, Zhenduo Zhang, ZuJie Wen +11
Training large reasoning models (LRMs) with reinforcement learning in STEM domains is hindered by the scarcity of high-quality, diverse, and verifiable problem sets. Existing synth…