long-context attention 1mixture-of-experts 1multimodal generation 1scaling laws 1visual diffusion models 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.CV2026
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers
Chongjian Ge, Hanwen Jiang, Tianyu Wang +9
The paper presents Chimera, a hybrid visual diffusion transformer that processes text, image, and video tokens in a single raster-ordered stream using efficient attention mechanism…
cs.CV2026
SimpleMatch: A Simple and Strong Baseline for Semantic Correspondence
Hailing Jin, Huiying Li
Recent advances in semantic correspondence have been largely driven by the use of pre-trained large-scale models. However, a limitation of these approaches is their dependence on h…