4 citations · 8 across the 2 of their papers we have counts for
6 papers
MACRO: Training-free Multi-plane Attention for Closeup Render Optimization
Nitzan Hodos, Roy Amoyal, Lior Fritz +3
Close-up rendering, zooming into a scene well beyond any training camera, is important for virtual production and interactive 3D content, yet remains an open challenge. 3D Gaussian…
SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPE
Or Hirschorn, Aaron Olender, Eli Alshan +3
We present a zero-shot, training-free and optimization-free framework for generating 360 panoramic images and videos by directly injecting spherical priors into pre-trained diffusi…
Splatent: Splatting Diffusion Latents for Novel View Synthesis
Or Hirschorn, Omer Sela, Inbar Huberman-Spiegelglas +6
Radiance field representations have recently been explored in the latent space of VAEs that are commonly used by diffusion models. This direction offers efficient rendering and sea…
Beyond Weak Perspective for Monocular 3D Human Pose Estimation
Imry Kissos, Lior Fritz, Matan Goldman +3
We consider the task of 3D joints location and orientation prediction from a monocular video with the skinned multi-person linear (SMPL) model. We first infer 2D joints locations w…
Joint Visual-Textual Embedding for Multimodal Style Search
Gil Sadeh, Lior Fritz, Gabi Shalev +1
We introduce a multimodal visual-textual search refinement method for fashion garments. Existing search engines do not enable intuitive, interactive, refinement of retrieved result…
Generating Diverse and Informative Natural Language Fashion Feedback
Gil Sadeh, Lior Fritz, Gabi Shalev +1
Recent advances in multi-modal vision and language tasks enable a new set of applications. In this paper, we consider the task of generating natural language fashion feedback on ou…