2 citations · 2 across the 7 of their papers we have counts for
Showing cs.GRShow all
2 papers · 1 filter
cs.GR2026
Sound Sparks Motion: Audio and Text Tuning for Video Editing
AmirHossein Naghi Razlighi, Aryan Mikaeili, Ali Mahdavi-Amiri +2
Motion-centric video editing remains difficult for large generative video models, which often respond well to appearance changes but struggle to produce specific, localized actions…
cs.GR2026
Untwisting RoPE: Frequency Control for Shared Attention in DiTs
Aryan Mikaeili, Or Patashnik, Andrea Tagliasacchi +2
Positional encodings are essential to transformer-based generative models, yet their behavior in multimodal and attention-sharing settings is not fully understood. In this work, we…