13 citations · 26 across the 13 of their papers we have counts for
13 papers
Mono4DEditor: Text-Driven 4D Scene Editing from Monocular Video via Point-Level Localization of Language-Embedded Gaussians
Jin-Chuan Shi, Chengye Su, Jiajun Wang +2
Editing 4D scenes reconstructed from monocular videos based on text prompts is a valuable yet challenging task with broad applications in content creation and virtual environments.…
VeasyGuide: Personalized Visual Guidance for Low-vision Learners on Instructor Actions in Presentation Videos
Yotam Sechayk, Ariel Shamir, Amy Pavel +1
Instructors often rely on visual actions such as pointing, marking, and sketching to convey information in educational presentation videos. These subtle visual cues often lack verb…
Unimodal Strategies in Density-Based Clustering
Oron Nir, Jay Tenenbaum, Ariel Shamir
Density-based clustering methods often surpass centroid-based counterparts, when addressing data with noise or arbitrary data distributions common in real-world problems. In this s…
LightLab: Controlling Light Sources in Images with Diffusion Models
Nadav Magar, Amir Hertz, Eric Tabellion +4
We present a simple, yet effective diffusion-based method for fine-grained, parametric control over light sources in an image. Existing relighting methods either rely on multiple i…
FontCLIP: A Semantic Typography Visual-Language Model for Multilingual Font Applications
Yuki Tatsukawa, I-Chao Shen, Anran Qi +3
Acquiring the desired font for various design tasks can be challenging and requires professional typographic knowledge. While previous font retrieval or generation works have allev…
VCR: Video representation for Contextual Retrieval
Oron Nir, Idan Vidra, Avi Neeman +2
Streamlining content discovery within media archives requires integrating advanced data representations and effective visualization techniques for clear communication of video topi…