354 citations · 470 across the 11 of their papers we have counts for
3 papers · 1 filter
Video-SwinUNet: Spatio-temporal Deep Learning Framework for VFSS Instance Segmentation
Chengxi Zeng, Xinyu Yang, David Smithard +3
This paper presents a deep learning framework for medical video segmentation. Convolution neural network (CNN) and transformer-based methods have achieved great milestones in medic…
Back to the Future: Cycle Encoding Prediction for Self-supervised Contrastive Video Representation Learning
Xinyu Yang, Majid Mirmehdi, Tilo Burghardt
In this paper we show that learning video feature spaces in which temporal cycles are maximally predictable benefits action classification. In particular, we propose a novel learni…
Great Ape Detection in Challenging Jungle Camera Trap Footage via Attention-Based Spatial and Temporal Feature Blending
Xinyu Yang, Majid Mirmehdi, Tilo Burghardt
We propose the first multi-frame video object detection framework trained to detect great apes. It is applicable to challenging camera trap footage in complex jungle environments a…