1 paper
Yiguang Yang, Jiankun Peng, Xiaoming Wang +2
Fast-WAM shows that video-action co-training improves control without generating future video at inference, making the representation from a single video diffusion Transformer forw…