2 papers
cs.LG2026
torchtune: PyTorch native post-training library
Mark Obozov, Maxime Griot, Joseph Cummings +8
Modern LLMs typically require multistage training pipelines to achieve strong downstream performance, with post-training serving as the main interface for adapting open-weight mode…
cs.LG2024
Lie Group Decompositions for Equivariant Neural Networks
Mircea Mironenco, Patrick Forré
Invariance and equivariance to geometrical transformations have proven to be very useful inductive biases when training (convolutional) neural network models, especially in the low…