Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
GRR-CoCa: Leveraging LLM Mechanisms in Multimodal Model Architectures
Jake R. Patock, Nicole Catherine Lewis, Kevin McCoy +3
State-of-the-art (SOTA) image and text generation models are multimodal models that have many similarities to large language models (LLMs). Despite achieving strong performances, l…
cs.CV2024
Using Skew to Assess the Quality of GAN-generated Image Features
Lorenzo Luzi, Helen Jenne, Ryan Murray +1
The rapid advancement of Generative Adversarial Networks (GANs) necessitates the need to robustly evaluate these models. Among the established evaluation criteria, the FréchetInce…
cs.CV2024
Boomerang: Local sampling on image manifolds using diffusion models
Lorenzo Luzi, Paul M Mayer, Josue Casco-Rodriguez +2
The inference stage of diffusion models can be seen as running a reverse-time diffusion stochastic differential equation, where samples from a Gaussian latent distribution are tran…