activity
20242026
collaborators

9 papers

cs.SD2026

Multi Codec Discrete Diffusion Model for Text Guided Speech Inpainting and Editing

Iftach Shoham, Tali Dror, Oren Gal +3

Speech recordings often contain missing, corrupted, or incorrect regions that must be reconstructed or modified without re-synthesizing the entire utterance. Speech inpainting rest…

cs.CL2026

Where Vision Becomes Text: Locating the OCR Routing Bottleneck in Vision-Language Models

Jonathan Steinberg, Oren Gal

Vision-language models (VLMs) can read text from images, but where does this optical character recognition (OCR) information enter the language processing stream? We investigate th…

cs.SD2026

Split and Conquer Partial Deepfake Speech

Inbal Rimon, Oren Gal, Haim Permuter

Partial deepfake speech detection requires identifying manipulated regions that may occur within short temporal portions of an otherwise bona fide utterance, making the task partic…

cs.SD2026

Token-Based Audio Inpainting via Discrete Diffusion

Tali Dror, Iftach Shoham, Moshe Buchris +4

Audio inpainting seeks to restore missing segments in degraded recordings. Previous diffusion-based methods exhibit impaired performance when the missing region is large. We introd…

cs.SD2025

Unmasking Deepfakes: Leveraging Augmentations and Features Variability for Deepfake Speech Detection

Inbal Rimon, Oren Gal, Haim Permuter

Deepfake speech detection presents a growing challenge as generative audio technologies continue to advance. We propose a hybrid training framework that advances detection performa…

cs.RO2025

Autonomous Oil Spill Response Through Liquid Neural Trajectory Modeling and Coordinated Marine Robotics

Hadas C. Kuzmenko, David Ehevich, Oren Gal

Marine oil spills pose grave environmental and economic risks, threatening marine ecosystems, coastlines, and dependent industries. Predicting and managing oil spill trajectories i…