activity
20172023
most citedAudio Tagging by Cross Filtering Noisy Labels

21 citations · 56 across the 14 of their papers we have counts for

collaborators
Showing cs.CVShow all

7 papers · 1 filter

cs.CV20231 cited

Cheap-fake Detection with LLM using Prompt Engineering

Guangyang Wu, Weijie Wu, Xiaohong Liu +3

The misuse of real photographs with conflicting image captions in news items is an example of the out-of-context (OOC) misuse of media. In order to detect OOC media, individuals mu…

cs.CV20221 cited

Trusted Multi-Scale Classification Framework for Whole Slide Image

Ming Feng, Kele Xu, Nanhui Wu +4

Despite remarkable efforts been made, the classification of gigapixels whole-slide image (WSI) is severely restrained from either the constrained computing resources for the whole…

cs.CV20214 cited

Multimodal Feature Fusion for Video Advertisements Tagging Via Stacking Ensemble

Qingsong Zhou, Hai Liang, Zhimin Lin +1

Automated tagging of video advertisements has been a critical yet challenging problem, and it has drawn increasing interests in last years as its applications seem to be evident in…

cs.CV20212 cited

Convolutional Neural Network-Based Age Estimation Using B-Mode Ultrasound Tongue Image

Kele Xu, Tamas Gábor Csapó, Ming Feng

Ultrasound tongue imaging is widely used for speech production research, and it has attracted increasing attention as its potential applications seem to be evident in many differen…

cs.CV202014 cited

AIM 2020: Scene Relighting and Illumination Estimation Challenge

Majed El Helou, Ruofan Zhou, Sabine Süsstrunk +34

We review the AIM 2020 challenge on virtual image relighting and illumination estimation. This paper presents the novel VIDIT dataset used in the challenge and the different propos…

cs.CV20193 cited

Predicting tongue motion in unlabeled ultrasound videos using convolutional LSTM neural network

Chaojie Zhao, Peng Zhang, Jian Zhu +3

A challenge in speech production research is to predict future tongue movements based on a short period of past tongue movements. This study tackles speaker-dependent tongue motion…