55 citations · 79 across the 4 of their papers we have counts for
5 papers
PromptSource: An Integrated Development Environment and Repository for Natural Language Prompts
Stephen H. Bach, Victor Sanh, Zheng-Xin Yong +24
PromptSource is a system for creating, sharing, and using natural language prompts. Prompts are functions that map an example from a dataset to a natural language input and target…
Masader: Metadata Sourcing for Arabic Text and Speech Data Resources
Zaid Alyafeai, Maraim Masoud, Mustafa Ghaleb +1
The NLP pipeline has evolved dramatically in the last few years. The first step in the pipeline is to find suitable annotated datasets to evaluate the tasks we are trying to solve.…
Calliar: An Online Handwritten Dataset for Arabic Calligraphy
Zaid Alyafeai, Maged S. Al-shaibani, Mustafa Ghaleb +1
Calligraphy is an essential part of the Arabic heritage and culture. It has been used in the past for the decoration of houses and mosques. Usually, such calligraphy is designed ma…
Evaluating Various Tokenizers for Arabic Text Classification
Zaid Alyafeai, Maged S. Al-shaibani, Mustafa Ghaleb +1
The first step in any NLP pipeline is to split the text into individual tokens. The most obvious and straightforward approach is to use words as tokens. However, given a large text…
A Survey on Transfer Learning in Natural Language Processing
Zaid Alyafeai, Maged Saeed AlShaibani, Irfan Ahmad
Deep learning models usually require a huge amount of data. However, these large datasets are not always attainable. This is common in many challenging NLP tasks. Consider Neural M…