1 paper
Wenya Yu, Chao Zhang, Li Wang +2
Post-Training Quantization (PTQ) and Low-Rank Adaptation (LoRA) constitute the standard pipeline for efficient Large Language Model (LLM) deployment. However, applying them sequent…