19 citations · 56 across the 19 of their papers we have counts for
4 papers · 2 filters
Analyzing Multilingual Competency of LLMs in Multi-Turn Instruction Following: A Case Study of Arabic
Sabri Boughorbel, Majd Hawasly
While significant progress has been made in benchmarking Large Language Models (LLMs) across various tasks, there is a lack of comprehensive evaluation of their abilities in respon…
Scaling up Discovery of Latent Concepts in Deep NLP Models
Majd Hawasly, Fahim Dalvi, Nadir Durrani
Despite the revolution caused by deep NLP models, they remain black boxes, necessitating research to understand their decision-making processes. A recent work by Dalvi et al. (2022…
LLMeBench: A Flexible Framework for Accelerating LLMs Benchmarking
Fahim Dalvi, Maram Hasanain, Sabri Boughorbel +10
The recent development and success of Large Language Models (LLMs) necessitate an evaluation of their performance across diverse NLP tasks in different languages. Although several…
LAraBench: Benchmarking Arabic AI with Large Language Models
Ahmed Abdelali, Hamdy Mubarak, Shammur Absar Chowdhury +13
Recent advancements in Large Language Models (LLMs) have significantly influenced the landscape of language and speech research. Despite this progress, these models lack specific b…