11 citations · 11 across the 3 of their papers we have counts for
3 papers
cs.CV2025
Analysis of Attention in Video Diffusion Transformers
Yuxin Wen, Jim Wu, Ajay Jain +2
We conduct an in-depth analysis of attention in video diffusion transformers (VDiTs) and report a number of novel findings. We identify three key properties of attention in VDiTs:…
cs.LG2023★ 11 cited
Video Prediction Models as Rewards for Reinforcement Learning
Alejandro Escontrela, Ademi Adeniji, Wilson Yan +6
Specifying reward signals that allow agents to learn complex behaviors is a long-standing challenge in reinforcement learning. A promising approach is to extract preferences for be…
cs.LG2022
AdaCat: Adaptive Categorical Discretization for Autoregressive Models
Qiyang Li, Ajay Jain, Pieter Abbeel
Autoregressive generative models can estimate complex continuous data distributions, like trajectory rollouts in an RL environment, image intensities, and audio. Most state-of-the-…