21 citations · 23 across the 2 of their papers we have counts for
4 papers
"Notic My Speech" -- Blending Speech Patterns With Multimedia
Dhruva Sahrawat, Yaman Kumar, Shashwat Aggarwal +3
Speech as a natural signal is composed of three parts - visemes (visual part of speech), phonemes (spoken part of speech), and language (the imposed structure). However, video as a…
Heterogeneity Loss to Handle Intersubject and Intrasubject Variability in Cancer
Shubham Goswami, Suril Mehta, Dhruva Sahrawat +2
Developing nations lack adequate number of hospitals with modern equipment and skilled doctors. Hence, a significant proportion of these nations' population, particularly in rural…
Keyphrase Extraction from Scholarly Articles as Sequence Labeling using Contextualized Embeddings
Dhruva Sahrawat, Debanjan Mahata, Mayank Kulkarni +7
In this paper, we formulate keyphrase extraction from scholarly articles as a sequence labeling task solved using a BiLSTM-CRF, where the words in the input text are represented us…
Harnessing GANs for Zero-shot Learning of New Classes in Visual Speech Recognition
Yaman Kumar, Dhruva Sahrawat, Shubham Maheshwari +5
Visual Speech Recognition (VSR) is the process of recognizing or interpreting speech by watching the lip movements of the speaker. Recent machine learning based approaches model VS…