2 papers
cs.AI2026
OPSD Compresses What RLVR Teaches: A Post-RL Compaction Stage for Reasoning Models
Jaehoon Kim, Dongha Lee
On-Policy Self-Distillation (OPSD) has recently emerged as an alternative to Reinforcement Learning with Verifiable Rewards (RLVR), promising higher accuracy and shorter responses…
math.DG2026
A Ruh-Vilms theorem for hypersurfaces in Weitzenböck geometry
Dongha Lee
A well-known theorem by Ruh and Vilms states that the Laplacian of the Gauss map for a smooth immersion into Euclidean space is given by the covariant derivative of the mean curvat…