444 citations · 1.7k across the 99 of their papers we have counts for
31 papers · 1 filter
I2MVFormer: Large Language Model Generated Multi-View Document Supervision for Zero-Shot Image Classification
Muhammad Ferjad Naeem, Muhammad Gul Zain Ali Khan, Yongqin Xian +4
Recent works have shown that unstructured text (documents) from online sources can serve as useful auxiliary information for zero-shot image classification. However, these methods…
Piecewise Planar Hulls for Semi-Supervised Learning of 3D Shape and Pose from 2D Images
Yigit Baran Can, Alexander Liniger, Danda Pani Paudel +1
We study the problem of estimating 3D shape and pose of an object in terms of keypoints, from a single 2D image. The shape and pose are learned directly from images collected by ca…
MicroISP: Processing 32MP Photos on Mobile Devices with Deep Learning
Andrey Ignatov, Anastasia Sycheva, Radu Timofte +8
While neural networks-based photo processing solutions can provide a better image quality compared to the traditional ISP systems, their application to mobile devices is still very…
PyNet-V2 Mobile: Efficient On-Device Photo Processing With Neural Networks
Andrey Ignatov, Grigory Malivenko, Radu Timofte +8
The increased importance of mobile photography created a need for fast and performant RAW image processing pipelines capable of producing good visual results in spite of the mobile…
Towards Versatile Embodied Navigation
Hanqing Wang, Wei Liang, Luc Van Gool +1
With the emergence of varied visual navigation tasks (e.g, image-/object-/audio-goal and vision-language navigation) that specify the target in different ways, the community has ma…
TripletTrack: 3D Object Tracking using Triplet Embeddings and LSTM
Nicola Marinello, Marc Proesmans, Luc Van Gool
3D object tracking is a critical task in autonomous driving systems. It plays an essential role for the system's awareness about the surrounding environment. At the same time there…