9 citations · 9 across the 3 of their papers we have counts for
3 papers
cs.CV2025
Synthetic Visual Genome
Jae Sung Park, Zixian Ma, Linjie Li +9
Reasoning over visual relationships-spatial, functional, interactional, social, etc.-is considered to be a fundamental component of human cognition. Yet, despite the major advances…
cs.CV2024
Iterated Learning Improves Compositionality in Large Vision-Language Models
Chenhao Zheng, Jieyu Zhang, Aniruddha Kembhavi +1
A fundamental characteristic common to both human vision and natural language is their compositional nature. Yet, despite the performance gains contributed by large vision and lang…
eess.IV2021★ 9 cited
A-ESRGAN: Training Real-World Blind Super-Resolution with Attention U-Net Discriminators
Zihao Wei, Yidong Huang, Yuang Chen +2
Blind image super-resolution(SR) is a long-standing task in CV that aims to restore low-resolution images suffering from unknown and complex distortions. Recent work has largely fo…