3 citations · 8 across the 13 of their papers we have counts for
8 papers · 1 filter
S2D: Sorted Speculative Decoding For More Efficient Deployment of Nested Large Language Models
Parsa Kavehzadeh, Mohammadreza Pourreza, Mojtaba Valipour +5
Deployment of autoregressive large language models (LLMs) is costly, and as these models increase in size, the associated costs will become even more considerable. Consequently, di…
EWEK-QA: Enhanced Web and Efficient Knowledge Graph Retrieval for Citation-based Question Answering Systems
Mohammad Dehghan, Mohammad Ali Alomrani, Sunyam Bagga +12
The emerging citation-based QA systems are gaining more attention especially in generative AI search applications. The importance of extracted knowledge provided to these systems i…
OTTAWA: Optimal TransporT Adaptive Word Aligner for Hallucination and Omission Translation Errors Detection
Chenyang Huang, Abbas Ghaddar, Ivan Kobyzev +3
Recently, there has been considerable attention on detecting hallucinations and omissions in Machine Translation (MT) systems. The two dominant approaches to tackle this task invol…
On the importance of Data Scale in Pretraining Arabic Language Models
Abbas Ghaddar, Philippe Langlais, Mehdi Rezagholizadeh +1
Pretraining monolingual language models have been proven to be vital for performance in Arabic Natural Language Processing (NLP) tasks. In this paper, we conduct a comprehensive st…
Mitigating Outlier Activations in Low-Precision Fine-Tuning of Language Models
Alireza Ghaffari, Justin Yu, Mahsa Ghazvini Nejad +3
Low-precision fine-tuning of language models has gained prominence as a cost-effective and energy-efficient approach to deploying large-scale models in various applications. Howeve…
Translate the Beauty in Songs: Jointly Learning to Align Melody and Translate Lyrics
Chengxi Li, Kai Fan, Jiajun Bu +3
Song translation requires both translation of lyrics and alignment of music notes so that the resulting verse can be sung to the accompanying melody, which is a challenging problem…