14 citations · 23 across the 8 of their papers we have counts for
7 papers · 1 filter
Zooming into Comics: Region-Aware RL Improves Fine-Grained Comic Understanding in Vision-Language Models
Yule Chen, Yufan Ren, Sabine Süsstrunk
Complex visual narratives, such as comics, present a significant challenge to Vision-Language Models (VLMs). Despite excelling on natural images, VLMs often struggle with stylized…
VibrantLeaves: A principled parametric image generator for training deep restoration models
Raphael Achddou, Yann Gousseau, Saïd Ladjal +1
In this paper, we introduce a synthetic image generator relying on a few simple principles, specifically focusing on geometric modeling, textures, and a simple modeling of image ac…
DSR: Towards Drone Image Super-Resolution
Xiaoyu Lin, Baran Ozaydin, Vidit Vidit +2
Despite achieving remarkable progress in recent years, single-image super-resolution methods are developed with several limitations. Specifically, they are trained on fixed content…
Volumetric Transformer Networks
Seungryong Kim, Sabine Süsstrunk, Mathieu Salzmann
Existing techniques to encode spatial invariance within deep convolutional neural networks (CNNs) apply the same warping field to all the feature channels. This does not account fo…
Editing in Style: Uncovering the Local Semantics of GANs
Edo Collins, Raja Bala, Bob Price +1
While the quality of GAN image synthesis has improved tremendously in recent years, our ability to control and condition the output is still limited. Focusing on StyleGAN, we intro…
Uniform Information Segmentation
Radhakrishna Achanta, Pablo Márquez-Neila, Pascal Fua +1
Size uniformity is one of the main criteria of superpixel methods. But size uniformity rarely conforms to the varying content of an image. The chosen size of the superpixels theref…