13 citations · 25 across the 3 of their papers we have counts for
3 papers
cs.LG2023★ 13 cited
Convolutional State Space Models for Long-Range Spatiotemporal Modeling
Jimmy T. H. Smith, Shalini De Mello, Jan Kautz +2
Effectively modeling long spatiotemporal sequences is challenging due to the need to model complex spatial correlations and long-range temporal dependencies simultaneously. ConvLST…
cs.SD2023★ 1 cited
The Power of Sound (TPoS): Audio Reactive Video Generation with Stable Diffusion
Yujin Jeong, Wonjeong Ryoo, Seunghyun Lee +4
In recent years, video generation has become a prominent generative tool and has drawn significant attention. However, there is little consideration in audio-to-video generation, t…
cs.CV2023★ 11 cited
Open-Vocabulary Panoptic Segmentation with Text-to-Image Diffusion Models
Jiarui Xu, Sifei Liu, Arash Vahdat +3
We present ODISE: Open-vocabulary DIffusion-based panoptic SEgmentation, which unifies pre-trained text-image diffusion and discriminative models to perform open-vocabulary panopti…