◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Pankaj Wasnik

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author1

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.CV2

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.CL2024

Efficient infusion of self-supervised representations in Automatic Speech Recognition

Darshan Prabhu, Sai Ganesh Mirishkar, Pankaj Wasnik

Self-supervised learned (SSL) models such as Wav2vec and HuBERT yield state-of-the-art results on speech-related tasks. Given the effectiveness of such models, it is advantageous t…

cs.CL2024

Isometric Neural Machine Translation using Phoneme Count Ratio Reward-based Reinforcement Learning

Shivam Ratnakant Mhaskar, Nirmesh J. Shah, Mohammadi Zaki +3

Traditional Automatic Video Dubbing (AVD) pipeline consists of three key modules, namely, Automatic Speech Recognition (ASR), Neural Machine Translation (NMT), and Text-to-Speech (…

cs.CV2024

Fiducial Focus Augmentation for Facial Landmark Detection

Purbayan Kar, Vishal Chudasama, Naoyuki Onoe +2

Deep learning methods have led to significant improvements in the performance on the facial landmark detection (FLD) task. However, detecting landmarks in challenging settings, suc…

cs.CV2023

Revisiting Class Imbalance for End-to-end Semi-Supervised Object Detection

Purbayan Kar, Vishal Chudasama, Naoyuki Onoe +1

Semi-supervised object detection (SSOD) has made significant progress with the development of pseudo-label-based end-to-end methods. However, many of these methods face challenges…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.