2 papers
cs.CV2025
AdaSVD: Adaptive Singular Value Decomposition for Large Language Models
Zhiteng Li, Mingyuan Xia, Jingyuan Zhang +5
Large language models (LLMs) have achieved remarkable success in natural language processing (NLP) tasks, yet their substantial memory requirements present significant challenges f…
cs.CV2025
QuantCache: Adaptive Importance-Guided Quantization with Hierarchical Latent and Layer Caching for Video Generation
Junyi Wu, Zhiteng Li, Zheng Hui +3
Recently, Diffusion Transformers (DiTs) have emerged as a dominant architecture in video generation, surpassing U-Net-based models in terms of performance. However, the enhanced ca…