3 papers
cs.AI2026
The Shape of Overthinking: Backtracking Bursts in Long Reasoning Traces
Navid Rezazadeh, Arash Gholami Davoodi
Reasoning models often generate long traces in which useful self-correction and unproductive revision are hard to distinguish. We study this distinction through backtracking dynami…
cs.CL2026
Geometry-Aware Decoding with Wasserstein-Regularized Truncation and Mass Penalties for Large Language Models
Arash Gholami Davoodi, Navid Rezazadeh, Seyed Pouyan Mousavi Davoudi +1
Large language models (LLMs) must balance diversity and creativity against logical coherence in open-ended generation. Existing truncation-based samplers are effective but largely…
cs.LG2026
Vertex-Softmax: Tight Transformer Verification via Exact Softmax Optimization
Navid Rezazadeh, Arash Gholami Davoodi
Certified verification of transformer attention requires bounding the softmax function over interval constraints on the pre-softmax scores. Existing verifiers relax softmax ndepend…