2 citations · 3 across the 4 of their papers we have counts for
4 papers
RealisVSR: Detail-enhanced Diffusion for Real-World 4K Video Super-Resolution
Weisong Zhao, Jingkai Zhou, Xiangyu Zhu +4
Video Super-Resolution (VSR) has achieved significant progress through diffusion models, effectively addressing the over-smoothing issues inherent in GAN-based methods. Despite rec…
Enhancing Instruction-Following Capability of Visual-Language Models by Reducing Image Redundancy
Te Yang, Jian Jia, Xiangyu Zhu +9
Large Language Models (LLMs) have strong instruction-following capability to interpret and execute tasks as directed by human commands. Multimodal Large Language Models (MLLMs) hav…
FRCSyn Challenge at WACV 2024:Face Recognition Challenge in the Era of Synthetic Data
Pietro Melzi, Ruben Tolosana, Ruben Vera-Rodriguez +44
Despite the widespread adoption of face recognition technology around the world, and its remarkable performance on current benchmarks, there are still several challenges that must…
Grouped Knowledge Distillation for Deep Face Recognition
Weisong Zhao, Xiangyu Zhu, Kaiwen Guo +2
Compared with the feature-based distillation methods, logits distillation can liberalize the requirements of consistent feature dimension between teacher and student networks, whil…