Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Continuous Speculative Decoding for Autoregressive Image Generation
Zili Wang, Zheng Zhang, Kun Ding +3
Continuous visual autoregressive (AR) models have demonstrated promising performance in image generation, but their inherently sequential nature results in slow inference speed. Sp…
cs.CV2024
AVESFormer: Efficient Transformer Design for Real-Time Audio-Visual Segmentation
Zili Wang, Qi Yang, Linsu Shi +4
Recently, transformer-based models have demonstrated remarkable performance on audio-visual segmentation (AVS) tasks. However, their expensive computational cost makes real-time in…