2 papers
cs.CL2024
Large Language Models Are Overparameterized Text Encoders
Thennal D K, Tim Fischer, Chris Biemann
Large language models (LLMs) demonstrate strong performance as text embedding models when finetuned with supervised contrastive training. However, their large size balloons inferen…
cs.CL2024
Advocating Character Error Rate for Multilingual ASR Evaluation
Thennal D K, Jesin James, Deepa P Gopinath +1
Automatic speech recognition (ASR) systems have traditionally been evaluated using English datasets, with the word error rate (WER) serving as the predominant metric. WER's simplic…