21 citations · 56 across the 14 of their papers we have counts for
7 papers · 1 filter
Cheap-fake Detection with LLM using Prompt Engineering
Guangyang Wu, Weijie Wu, Xiaohong Liu +3
The misuse of real photographs with conflicting image captions in news items is an example of the out-of-context (OOC) misuse of media. In order to detect OOC media, individuals mu…
Trusted Multi-Scale Classification Framework for Whole Slide Image
Ming Feng, Kele Xu, Nanhui Wu +4
Despite remarkable efforts been made, the classification of gigapixels whole-slide image (WSI) is severely restrained from either the constrained computing resources for the whole…
Multimodal Feature Fusion for Video Advertisements Tagging Via Stacking Ensemble
Qingsong Zhou, Hai Liang, Zhimin Lin +1
Automated tagging of video advertisements has been a critical yet challenging problem, and it has drawn increasing interests in last years as its applications seem to be evident in…
Convolutional Neural Network-Based Age Estimation Using B-Mode Ultrasound Tongue Image
Kele Xu, Tamas Gábor Csapó, Ming Feng
Ultrasound tongue imaging is widely used for speech production research, and it has attracted increasing attention as its potential applications seem to be evident in many differen…
AIM 2020: Scene Relighting and Illumination Estimation Challenge
Majed El Helou, Ruofan Zhou, Sabine Süsstrunk +34
We review the AIM 2020 challenge on virtual image relighting and illumination estimation. This paper presents the novel VIDIT dataset used in the challenge and the different propos…
Predicting tongue motion in unlabeled ultrasound videos using convolutional LSTM neural network
Chaojie Zhao, Peng Zhang, Jian Zhu +3
A challenge in speech production research is to predict future tongue movements based on a short period of past tongue movements. This study tackles speaker-dependent tongue motion…