1 paper
Junhao Wu, Dezhong Yao, Hai Jin
W4A4 quantization of large video diffusion Transformers offers substantial memory savings but is hindered by two main challenges: sparse large-magnitude activation outliers, and st…