20 papers
MV-Forcing: Long Multi-View Video Generation via 4D-Grounded Spatio-Temporal Self-Forcing
Gal Fiebelman, Hadar Averbuch-Elor, Sagie Benaim
Recent advances in video diffusion models have enabled either long single-view generation through temporal autoregression, or short multi-view synthesis through bidirectional atten…
MACRO: Training-free Multi-plane Attention for Closeup Render Optimization
Nitzan Hodos, Roy Amoyal, Lior Fritz +3
Close-up rendering, zooming into a scene well beyond any training camera, is important for virtual production and interactive 3D content, yet remains an open challenge. 3D Gaussian…
TrajLoc: Trajectory-Attention Localization for Multi-Object Motion Control
Omer Sela, Inbar Huberman-Spiegelglas, Michael Rotman +2
Controlling the motion of multiple objects in image-to-video (I2V) generation requires preserving object identities while enforcing adherence to distinct target trajectories. This…
SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE
Or Hirschorn, Aaron Olender, Eli Alshan +3
We present a zero-shot, training-free and optimization-free framework for generating 360 panoramic images and videos by directly injecting spherical priors into pre-trained diffusi…
Colored Noise Diffusion Sampling
Hadar Davidson, Noam Issachar, Sagie Benaim
Diffusion models achieve state-of-the-art image synthesis, with their generative trajectories fundamentally exhibiting a spectral bias, resolving low-frequency global structures ea…
PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions
Omer Benishu, Gal Fiebelman, Sagie Benaim
We address the task of generating physically accurate and visually faithful 4D Human-Object Interaction (HOI). Given a static 3D human and target object represented as 3D Gaussian…