32 citations · 32 across the 2 of their papers we have counts for
2 papers
cs.CV2025
Pushing the Boundaries of State Space Models for Image and Video Generation
Yicong Hong, Long Mai, Yuan Yao +1
While Transformers have become the dominant architecture for visual generation, linear attention models, such as the state-space models (SSM), are increasingly recognized for their…
cs.CV2023★ 32 cited
Instant3D: Fast Text-to-3D with Sparse-View Generation and Large Reconstruction Model
Jiahao Li, Hao Tan, Kai Zhang +7
Text-to-3D with diffusion models has achieved remarkable progress in recent years. However, existing methods either rely on score distillation-based optimization which suffer from…