1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2023
Exploring Compositional Visual Generation with Latent Classifier Guidance
Changhao Shi, Haomiao Ni, Kai Li +3
Diffusion probabilistic models have achieved enormous success in the field of image generation and manipulation. In this paper, we explore a novel paradigm of using the diffusion m…
cs.CV2022
Learning Transferable Reward for Query Object Localization with Policy Adaptation
Tingfeng Li, Shaobo Han, Martin Renqiang Min +1
We propose a reinforcement learning based approach to query object localization, for which an agent is trained to localize objects of interest specified by a small exemplary set. W…
cs.CV2021★ 1 cited
Hopper: Multi-hop Transformer for Spatiotemporal Reasoning
Honglu Zhou, Asim Kadav, Farley Lai +4
This paper considers the problem of spatiotemporal object-centric reasoning in videos. Central to our approach is the notion of object permanence, i.e., the ability to reason about…