40 citations · 40 across the 3 of their papers we have counts for
3 papers
cs.CV2025
Analysis of Attention in Video Diffusion Transformers
Yuxin Wen, Jim Wu, Ajay Jain +2
We conduct an in-depth analysis of attention in video diffusion transformers (VDiTs) and report a number of novel findings. We identify three key properties of attention in VDiTs:…
cs.CV2024
EditScout: Locating Forged Regions from Diffusion-based Edited Images with Multimodal LLM
Quang Nguyen, Truong Vu, Trong-Tung Nguyen +6
Image editing technologies are tools used to transform, adjust, remove, or otherwise alter images. Recent research has significantly improved the capabilities of image editing tool…
cs.LG2023★ 40 cited
Hard Prompts Made Easy: Gradient-Based Discrete Optimization for Prompt Tuning and Discovery
Yuxin Wen, Neel Jain, John Kirchenbauer +3
The strength of modern generative models lies in their ability to be controlled through text-based prompts. Typical "hard" prompts are made from interpretable words and tokens, and…