Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Stable Adaptive Thinking via Advantage Shaping and Length-Aware Gradient Regulation
Zihang Xu, Haozhi Xie, Ziqi Miao +3
Large reasoning models (LRMs) achieve strong performance through extended reasoning traces, but they often exhibit overthinking behavior for low-complexity queries. Existing effort…
cs.LG2026
TabSieve: Explicit In-Table Evidence Selection for Tabular Prediction
Yongyao Wang, Ziqi Miao, Lu Yang +4
Tabular prediction can benefit from in-table rows as few-shot evidence, yet existing tabular models typically perform instance-wise inference and LLM-based prompting is often britt…