3 papers
cs.CV2026
PrivLEX: Detecting legal concepts in images through Vision-Language Models
Darya Baranouskaya, Andrea Cavallaro
We present PrivLEX, a novel image privacy classifier that grounds its decisions in legally defined personal data concepts. PrivLEX is the first interpretable privacy classifier ali…
cs.SD2025
Shortcut Flow Matching for Speech Enhancement: Step-Invariant flows via single stage training
Naisong Zhou, Saisamarth Rajesh Phaye, Milos Cernak +4
Diffusion-based generative models have achieved state-of-the-art performance for perceptual quality in speech enhancement (SE). However, their iterative nature requires numerous Ne…
cs.CV2025
3D Face Reconstruction Error Decomposed: A Modular Benchmark for Fair and Fast Method Evaluation
Evangelos Sariyanidi, Claudio Ferrari, Federico Nocentini +3
Computing the standard benchmark metric for 3D face reconstruction, namely geometric error, requires a number of steps, such as mesh cropping, rigid alignment, or point corresponde…