1 citations · 1 across the 2 of their papers we have counts for
3 papers · 1 filter
Timbre-Aware LLM-based Direct Speech-to-Speech Translation Extendable to Multiple Language Pairs
Lalaram Arya, Mrinmoy Bhattacharjee, Adarsh C. R. +1
Direct Speech-to-Speech Translation (S2ST) has gained increasing attention for its ability to translate speech from one language to another, while reducing error propagation and la…
Fusion of Modulation Spectrogram and SSL with Multi-head Attention for Fake Speech Detection
Rishith Sadashiv T N, Abhishek Bedge, Saisha Suresh Bore +3
Fake speech detection systems have become a necessity to combat against speech deepfakes. Current systems exhibit poor generalizability on out-of-domain speech samples due to lack…
Time-Frequency Audio Features for Speech-Music Classification
Mrinmoy Bhattacharjee, S. R. M. Prasanna, Prithwijit Guha
Distinct striation patterns are observed in the spectrograms of speech and music. This motivated us to propose three novel time-frequency features for speech-music classification.…