15 papers
Multi Codec Discrete Diffusion Model for Text Guided Speech Inpainting and Editing
Iftach Shoham, Tali Dror, Oren Gal +3
Speech recordings often contain missing, corrupted, or incorrect regions that must be reconstructed or modified without re-synthesizing the entire utterance. Speech inpainting rest…
Plan for Speed: Dilated Scheduling for Masked Diffusion Language Models
Omer Luxembourg, Haim Permuter, Eliya Nachmani
Masked diffusion language models (MDLMs) promise fast, non-autoregressive text generation, yet existing samplers, which pick tokens to unmask based on model confidence, ignore inte…
Split and Conquer Partial Deepfake Speech
Inbal Rimon, Oren Gal, Haim Permuter
Partial deepfake speech detection requires identifying manipulated regions that may occur within short temporal portions of an otherwise bona fide utterance, making the task partic…
Token-Based Audio Inpainting via Discrete Diffusion
Tali Dror, Iftach Shoham, Moshe Buchris +4
Audio inpainting seeks to restore missing segments in degraded recordings. Previous diffusion-based methods exhibit impaired performance when the missing region is large. We introd…
Directed Information: Estimation, Optimization and Applications in Communications and Causality
Dor Tsur, Oron Sabag, Navin Kashyap +2
Directed information (DI) is an information measure that attempts to capture directionality in the flow of information from one random process to another. It is closely related to…
HeavyWater and SimplexWater: Distortion-Free LLM Watermarks for Low-Entropy Next-Token Predictions
Dor Tsur, Carol Xuan Long, Claudio Mayrink Verdun +5
Large language model (LLM) watermarks enable authentication of text provenance, curb misuse of machine-generated text, and promote trust in AI systems. Current watermarks operate b…