11 citations · 24 across the 19 of their papers we have counts for
14 papers · 1 filter
Giving a Hand to Diffusion Models: a Two-Stage Approach to Improving Conditional Human Image Generation
Anton Pelykh, Ozge Mercanoglu Sincan, Richard Bowden
Recent years have seen significant progress in human image generation, particularly with the advancements in diffusion models. However, existing diffusion methods encounter challen…
The Third Monocular Depth Estimation Challenge
Jaime Spencer, Fabio Tosi, Matteo Poggi +38
This paper discusses the results of the third edition of the Monocular Depth Estimation Challenge (MDEC). The challenge focuses on zero-shot generalization to the challenging SYNS-…
Two Hands Are Better Than One: Resolving Hand to Hand Intersections via Occupancy Networks
Maksym Ivashechkin, Oscar Mendez, Richard Bowden
3D hand pose estimation from images has seen considerable interest from the literature, with new methods improving overall 3D accuracy. One current challenge is to address hand-to-…
Kick Back & Relax++: Scaling Beyond Ground-Truth Depth with SlowTV & CribsTV
Jaime Spencer, Chris Russell, Simon Hadfield +1
Self-supervised learning is the key to unlocking generic computer vision systems. By eliminating the reliance on ground-truth annotations, it allows scaling to much larger data qua…
Is context all you need? Scaling Neural Sign Language Translation to Large Domains of Discourse
Ozge Mercanoglu Sincan, Necati Cihan Camgoz, Richard Bowden
Sign Language Translation (SLT) is a challenging task that aims to generate spoken language sentences from sign language videos, both of which have different grammar and word/gloss…
Improving 3D Pose Estimation for Sign Language
Maksym Ivashechkin, Oscar Mendez, Richard Bowden
This work addresses 3D human pose reconstruction in single images. We present a method that combines Forward Kinematics (FK) with neural networks to ensure a fast and valid predict…