37 citations · 37 across the 1 of their papers we have counts for
2 papers
cs.CV2023
EVP: Enhanced Visual Perception using Inverse Multi-Attentive Feature Refinement and Regularized Image-Text Alignment
Mykola Lavreniuk, Shariq Farooq Bhat, Matthias Müller +1
This work presents the network architecture EVP (Enhanced Visual Perception). EVP builds on the previous work VPD which paved the way to use the Stable Diffusion network for comput…
cs.CV2023★ 37 cited
MiDaS v3.1 -- A Model Zoo for Robust Monocular Relative Depth Estimation
Reiner Birkl, Diana Wofk, Matthias Müller
We release MiDaS v3.1 for monocular depth estimation, offering a variety of new models based on different encoder backbones. This release is motivated by the success of transformer…