5 papers
Train Short, Inference Long: Training-free Horizon Extension for Autoregressive Video Generation
Jia Li, Xiaomeng Fu, Xurui Peng +7
Autoregressive video diffusion models have emerged as a scalable paradigm for long video generation. However, they often suffer from severe extrapolation failure, where rapid error…
Contrastive Conditional-Unconditional Alignment for Long-tailed Diffusion Model
Fang Chen, Alex Villa, Gongbo Liang +3
Training data for class-conditional image synthesis often exhibit a long-tailed distribution with limited amount of images for tail classes. Such an imbalance causes mode collapse…
Efficient Masked Image Compression with Position-Indexed Self-Attention
Chengjie Dai, Tiantian Song, Hui Tang +3
In recent years, image compression for high-level vision tasks has attracted considerable attention from researchers. Given that object information in images plays a far more cruci…
ReDistill: Residual Encoded Distillation for Peak Memory Reduction of CNNs
Fang Chen, Gourav Datta, Mujahid Al Rafi +2
The expansion of neural network sizes and the enhanced resolution of modern image sensors result in heightened memory and power demands to process modern computer vision models. In…
SimQ-NAS: Simultaneous Quantization Policy and Neural Architecture Search
Sharath Nittur Sridhar, Maciej Szankin, Fang Chen +2
Recent one-shot Neural Architecture Search algorithms rely on training a hardware-agnostic super-network tailored to a specific task and then extracting efficient sub-networks for…