3 citations · 5 across the 3 of their papers we have counts for
5 papers · 1 filter
Éclair -- Extracting Content and Layout with Integrated Reading Order for Documents
Ilia Karmanov, Amala Sanjay Deshmukh, Lukas Voegtle +8
Optical Character Recognition (OCR) technology is widely used to extract text from images of documents, facilitating efficient digitization and data retrieval. However, merely extr…
Self-attention fusion for audiovisual emotion recognition with incomplete data
Kateryna Chumachenko, Alexandros Iosifidis, Moncef Gabbouj
In this paper, we consider the problem of multimodal data analysis with a use case of audiovisual emotion recognition. We propose an architecture capable of learning from raw data…
Self-Attention Neural Bag-of-Features
Kateryna Chumachenko, Alexandros Iosifidis, Moncef Gabbouj
In this work, we propose several attention formulations for multivariate sequence data. We build on top of the recently introduced 2D-Attention and reformulate the attention learni…
Ensembling object detectors for image and video data analysis
Kateryna Chumachenko, Jenni Raitoharju, Alexandros Iosifidis +1
In this paper, we propose a method for ensembling the outputs of multiple object detectors for improving detection performance and precision of bounding boxes on image data. We fur…
Machine Learning Based Analysis of Finnish World War II Photographers
Kateryna Chumachenko, Anssi Männistö, Alexandros Iosifidis +1
In this paper, we demonstrate the benefits of using state-of-the-art machine learning methods in the analysis of historical photo archives. Specifically, we analyze prominent Finni…