1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 1 cited
LinVT: Empower Your Image-level Large Language Model to Understand Videos
Lishuai Gao, Yujie Zhong, Yingsen Zeng +3
Large Language Models (LLMs) have been widely used in various tasks, motivating us to develop an LLM-based assistant for videos. Instead of training from scratch, we propose a modu…
cs.CV2024
RFSR: Improving ISR Diffusion Models via Reward Feedback Learning
Xiaopeng Sun, Qinwei Lin, Yu Gao +6
Generative diffusion models (DM) have been extensively utilized in image super-resolution (ISR). Most of the existing methods adopt the denoising loss from DDPMs for model optimiza…