51 citations · 70 across the 14 of their papers we have counts for
8 papers · 1 filter
Association Graph Learning for Multi-Task Classification with Category Shifts
Jiayi Shen, Zehao Xiao, Xiantong Zhen +2
In this paper, we focus on multi-task classification, where related classification tasks share the same label space and are learned simultaneously. In particular, we tackle a new s…
Longer Version for "Deep Context-Encoding Network for Retinal Image Captioning"
Jia-Hong Huang, Ting-Wei Wu, Chao-Han Huck Yang +1
Automatically generating medical reports for retinal images is one of the promising ways to help ophthalmologists reduce their workload and improve work efficiency. In this work, w…
Contextualized Keyword Representations for Multi-modal Retinal Image Captioning
Jia-Hong Huang, Ting-Wei Wu, Marcel Worring
Medical image captioning automatically generates a medical description to describe the content of a given medical image. A traditional medical image captioning model creates a medi…
GPT2MVS: Generative Pre-trained Transformer-2 for Multi-modal Video Summarization
Jia-Hong Huang, Luka Murn, Marta Mrak +1
Traditional video summarization methods generate fixed video representations regardless of user interest. Therefore such methods limit users' expectations in content search and exp…
Detecting CNN-Generated Facial Images in Real-World Scenarios
Nils Hulzebosch, Sarah Ibrahimi, Marcel Worring
Artificial, CNN-generated images are now of such high quality that humans have trouble distinguishing them from real images. Several algorithmic detection methods have been propose…
4-Connected Shift Residual Networks
Andrew Brown, Pascal Mettes, Marcel Worring
The shift operation was recently introduced as an alternative to spatial convolutions. The operation moves subsets of activations horizontally and/or vertically. Spatial convolutio…