1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.GR2025
Image Editing with Diffusion Models: A Survey
Jia Wang, Jie Hu, Xiaoqi Ma +3
With deeper exploration of diffusion model, developments in the field of image generation have triggered a boom in image creation. As the quality of base-model generated images con…
cs.LG2024
Denoising with a Joint-Embedding Predictive Architecture
Dengsheng Chen, Jie Hu, Xiaoming Wei +1
Joint-embedding predictive architectures (JEPAs) have shown substantial promise in self-supervised representation learning, yet their application in generative modeling remains und…
cs.CV2024★ 1 cited
Kangaroo: A Powerful Video-Language Model Supporting Long-context Video Input
Jiajun Liu, Yibing Wang, Hanghang Ma +6
Rapid advancements have been made in extending Large Language Models (LLMs) to Large Multi-modal Models (LMMs). However, extending input modality of LLMs to video data remains a ch…