3 citations · 7 across the 9 of their papers we have counts for
9 papers
ChimpVLM: Ethogram-Enhanced Chimpanzee Behaviour Recognition
Otto Brookes, Majid Mirmehdi, Hjalmar Kuhl +1
We show that chimpanzee behaviour understanding from camera traps can be enhanced by providing visual architectures with access to an embedding of text descriptions that detail spe…
PanAf20K: A Large Video Dataset for Wild Ape Detection and Behaviour Recognition
Otto Brookes, Majid Mirmehdi, Colleen Stephens +24
We present the PanAf20K dataset, the largest and most diverse open-access annotated video dataset of great apes in their natural environment. It comprises more than 7 million frame…
PECoP: Parameter Efficient Continual Pretraining for Action Quality Assessment
Amirhossein Dadashzadeh, Shuchao Duan, Alan Whone +1
The limited availability of labelled data in Action Quality Assessment (AQA), has forced previous works to fine-tune their models pretrained on large-scale domain-general datasets.…
Use Your Head: Improving Long-Tail Video Recognition
Toby Perrett, Saptarshi Sinha, Tilo Burghardt +2
This paper presents an investigation into long-tail video recognition. We demonstrate that, unlike naturally-collected video datasets and existing long-tail image benchmarks, curre…
Video-TransUNet: Temporally Blended Vision Transformer for CT VFSS Instance Segmentation
Chengxi Zeng, Xinyu Yang, Majid Mirmehdi +2
We propose Video-TransUNet, a deep architecture for instance segmentation in medical CT videos constructed by integrating temporal feature blending into the TransUNet deep learning…
Detecting Humans in RGB-D Data with CNNs
Kaiyang Zhou, Adeline Paiement, Majid Mirmehdi
We address the problem of people detection in RGB-D data where we leverage depth information to develop a region-of-interest (ROI) selection method that provides proposals to two c…