2 papers
cs.LG2026
On Repulsive and Attractive Teachers: Separating Correctness from Behavior in Self-Distillation
Anton Baumann, Akmal Ashirmatov, Leo Schmidt-Traub +4
On-policy self-distillation provides dense, token-level supervision by conditioning a model on privileged information and distilling the resulting teacher distribution back into th…
cs.SD2026
Geometric Iterative Retrieval for Neural Audio Codec Resynthesis
Leo Schmidt-Traub, Frédéric Berdoz, Luca A. Lanzendörfer +1
Neural audio codecs based on Residual Vector Quantization (RVQ) have become the dominant discrete representation for token-based general audio generation, yet resynthesizing high-q…