2 papers
cs.CV2025
Towards Fine-Grained Human Motion Video Captioning
Guorui Song, Guocun Wang, Zhe Huang +4
Generating accurate descriptions of human actions in videos remains a challenging task for video captioning models. Existing approaches often struggle to capture fine-grained motio…
cs.CV2025
TASR: Timestep-Aware Diffusion Model for Image Super-Resolution
Qinwei Lin, Xiaopeng Sun, Yu Gao +4
Diffusion models have recently achieved outstanding results in the field of image super-resolution. These methods typically inject low-resolution (LR) images via ControlNet.In this…