2 citations · 3 across the 4 of their papers we have counts for
12 papers
Image Captioning with Integrated Bottom-Up and Multi-level Residual Top-Down Attention for Game Scene Understanding
Jian Zheng, Sudha Krishnamurthy, Ruxin Chen +3
Image captioning has attracted considerable attention in recent years. However, little work has been done for game image captioning which has some unique characteristics and requir…
Neural Network Based Speaker Classification and Verification Systems with Enhanced Features
Zhenhao Ge, Ananth N. Iyer, Srinath Cheluvaraja +2
This work presents a novel framework based on feed-forward neural network for text-independent speaker classification and verification, two related systems of speaker recognition.…
Speaker Change Detection Using Features through A Neural Network Speaker Classifier
Zhenhao Ge, Ananth N. Iyer, Srinath Cheluvaraja +1
The mechanism proposed here is for real-time speaker change detection in conversations, which firstly trains a neural network text-independent speaker classifier using in-domain sp…
Generation and Pruning of Pronunciation Variants to Improve ASR Accuracy
Zhenhao Ge, Aravind Ganapathiraju, Ananth N. Iyer +2
Speech recognition, especially name recognition, is widely used in phone services such as company directory dialers, stock quote providers or location finders. It is usually challe…
Accent Classification with Phonetic Vowel Representation
Zhenhao Ge, Yingyi Tan, Aravind Ganapathiraju
Previous accent classification research focused mainly on detecting accents with pure acoustic information without recognizing accented speech. This work combines phonetic knowledg…
Adaptive Frequency Cepstral Coefficients for Word Mispronunciation Detection
Zhenhao Ge, Sudhendu R. Sharma, Mark J. T. Smith
Systems based on automatic speech recognition (ASR) technology can provide important functionality in computer assisted language learning applications. This is a young but growing…