2 papers
cs.CV2026
Order Matters: LVLMs as Judges for Temporal Reasoning in Image Sequences
Martina Ianaro, Guilherme Fernandes, Maurizio Gabbrielli +1
As generative multimedia evolves from static image synthesis to complex, interleaved visual narratives, a foundational bottleneck has emerged: the judgment crisis. While human perc…
cs.CV2025
Latent Beam Diffusion Models for Generating Visual Sequences
Guilherme Fernandes, Vasco Ramos, Regev Cohen +2
While diffusion models excel at generating high-quality images from text prompts, they struggle with visual consistency when generating image sequences. Existing methods generate e…