2 papers
cs.CV2025
Video Object Segmentation-Aware Audio Generation
Ilpo Viertola, Vladimir Iashin, Esa Rahtu
Existing multimodal audio generation models often lack precise user control, which limits their applicability in professional Foley workflows. In particular, these models focus on…
cs.CV2025
UDGS-SLAM : UniDepth Assisted Gaussian Splatting for Monocular SLAM
Mostafa Mansour, Ahmed Abdelsalam, Ari Happonen +2
Recent advancements in monocular neural depth estimation, particularly those achieved by the UniDepth network, have prompted the investigation of integrating UniDepth within a Gaus…