28 citations · 31 across the 2 of their papers we have counts for
5 papers · 1 filter
Feedback-Driven Vision-Language Alignment with Minimal Human Supervision
Giorgio Giannone, Ruoteng Li, Qianli Feng +3
Vision-language models (VLMs) have demonstrated remarkable potential in integrating visual and linguistic information, but their performance is often constrained by the need for ex…
AdAM: Few-Shot Image Generation via Adaptation-Aware Kernel Modulation
Yunqing Zhao, Keshigeyan Chandrasegaran, Milad Abdollahzadeh +5
Few-shot image generation (FSIG) aims to learn to generate new and diverse images given few (e.g., 10) training samples. Recent work has addressed FSIG by leveraging a GAN pre-trai…
Object Tracking Using Spatio-Temporal Future Prediction
Yuan Liu, Ruoteng Li, Robby T. Tan +2
Occlusion is a long-standing problem that causes many modern tracking methods to be erroneous. In this paper, we address the occlusion problem by exploiting the current and future…
Heavy Rain Image Restoration: Integrating Physics Model and Conditional Adversarial Learning
Ruotent Li, Loong Fah Cheong, Robby T. Tan
Most deraining works focus on rain streaks removal but they cannot deal adequately with heavy rain images. In heavy rain, streaks are strongly visible, dense rain accumulation or r…
Single Image Deraining using Scale-Aware Multi-Stage Recurrent Network
Ruoteng Li, Loong-Fah Cheong, Robby T. Tan
Given a single input rainy image, our goal is to visually remove rain streaks and the veiling effect caused by scattering and transmission of rain streaks and rain droplets. We are…