papers
Publications (2)
cs.CV2024
Imagen 3
Imagen-Team-Google, :, Jason Baldridge +257
We introduce Imagen 3, a latent diffusion model that generates high quality images from text prompts. We describe our quality and responsibility evaluations. Imagen 3 is preferred…
cs.CV2018
Activity Recognition on a Large Scale in Short Videos - Moments in Time Dataset
Ankit Shah, Harini Kesavamoorthy, Poorva Rane +3
Moments capture a huge part of our lives. Accurate recognition of these moments is challenging due to the diverse and complex interpretation of the moments. Action recognition refe…