2 papers
cs.CV2025
Punching Above Precision: Small Quantized Model Distillation with Learnable Regularizer
Abdur Rehman, S M A Sharif, Md Abdur Rahaman +3
Quantization-aware training (QAT) combined with knowledge distillation (KD) is a promising strategy for compressing Artificial Intelligence (AI) models for deployment on resource-c…
cs.LG2025
Riemannian Optimization for LoRA on the Stiefel Manifold
Juneyoung Park, Minjae Kang, Seongbae Lee +3
While powerful, large language models (LLMs) present significant fine-tuning challenges due to their size. Parameter-efficient fine-tuning (PEFT) methods like LoRA provide solution…