7 papers
Pixel-Space Diffusion via Observation Operators
Shaojie Guo, Lichen Ma, Haoyang Tong +8
Pixel-space diffusion models directly model image distributions but remain difficult to optimize. Recent methods alleviate this challenge through target reparameterization, while s…
LoViF 2026 The First Challenge on Unified Removal of Raindrops and Reflections: Methods and Results
Zewei He, Xi Tong, Yu Chen +49
This workshop paper comprehensively reviews the First Challenge on Unified Removal of Raindrops and Reflections. The challenge aims to address a frequently encountered practical pr…
GRNEdit: Efficient General Video Editing from a New Binary-Evidence Perspective in Generative Refinement Networks
Feng Xie, Jiagao Hu, Fuhao Li +5
Instruction-based general video editing seeks to unify diverse editing operations within a single, intuitive interface. Existing approaches often rely on resource-intensive conditi…
HapticLDM: A Diffusion Model for Text-to-Vibrotactile Generation
Jiahao Xiong, Fei Wang, Anran Xu +4
Text-to-vibration generation converts natural language into haptic feedback, enabling vibration-effect designers to get scenarios-fitted vibrations more efficiently, which shows gr…
PROVE: A Perceptual RemOVal cohErence Benchmark for Visual Media
Fuhao Li, Shaofeng You, Jiagao Hu +6
Evaluating object removal in images and videos remains challenging because the task is inherently one-to-many, yet existing metrics frequently disagree with human perception. Full-…
End4: End-to-end Denoising Diffusion for Diffusion-Based Inpainting Detection
Fei Wang, Xuecheng Wu, Zheng Zhang +3
The powerful generative capabilities of diffusion models have significantly advanced the field of image synthesis, enhancing both full image generation and inpainting-based image e…