ArCorona: Analyzing Arabic Tweets in the Early Days of Coronavirus (COVID-19) Pandemic
arXiv:2012.01462
Abstract
Over the past few months, there were huge numbers of circulating tweets and discussions about Coronavirus (COVID-19) in the Arab region. It is important for policy makers and many people to identify types of shared tweets to better understand public behavior, topics of interest, requests from governments, sources of tweets, etc. It is also crucial to prevent spreading of rumors and misinformation about the virus or bad cures. To this end, we present the largest manually annotated dataset of Arabic tweets related to COVID-19. We describe annotation guidelines, analyze our dataset and build effective machine learning and transformer based models for classification.
References in corpus (7)
- A large-scale COVID-19 Twitter chatter dataset for open scientific research -- an international collaboration
- ktrain: A Low-Code Library for Augmented Machine Learning
- Disinformation and Misinformation on Twitter during the Novel Coronavirus Outbreak
- ArCOV-19: The First Arabic COVID-19 Twitter Dataset with Propagation Networks
- Large Arabic Twitter Dataset on COVID-19
- Arabic Dialect Identification in the Wild
- Analysis of misinformation during the COVID-19 outbreak in China: cultural, social and political entanglements