1 citations · 1 across the 3 of their papers we have counts for
2 papers
cs.SD2023
VANI: Very-lightweight Accent-controllable TTS for Native and Non-native speakers with Identity Preservation
Rohan Badlani, Akshit Arora, Subhankar Ghosh +5
We introduce VANI, a very lightweight multi-lingual accent controllable speech synthesis system. Our model builds upon disentanglement strategies proposed in RADMMM and supports ex…
cs.IR2014
Efficient Media Retrieval from Non-Cooperative Queries
Kevin Shih, Wei Di, Vignesh Jagadeesh +1
Text is ubiquitous in the artificial world and easily attainable when it comes to book title and author names. Using the images from the book cover set from the Stanford Mobile Vis…