activity
20152021
most citedIndoLEM and IndoBERT: A Benchmark Dataset and Pre-trained Language Model for Indonesian NLP

32 citations · 37 across the 4 of their papers we have counts for

collaborators

8 papers

cs.CL2021

Fairness-aware Class Imbalanced Learning

Shivashankar Subramanian, Afshin Rahimi, Timothy Baldwin +2

Class imbalance is a common challenge in many NLP tasks, and has clear connections to bias, in that bias in training data often leads to higher accuracy for majority groups at the…

cs.CL20204 cited

Learning Causal Bayesian Networks from Text

Farhad Moghimifar, Afshin Rahimi, Mahsa Baktashmotlagh +1

Causal relationships form the basis for reasoning and decision-making in Artificial Intelligence systems. To exploit the large volume of textual data available today, the automatic…

cs.CL202032 cited

IndoLEM and IndoBERT: A Benchmark Dataset and Pre-trained Language Model for Indonesian NLP

Fajri Koto, Afshin Rahimi, Jey Han Lau +1

Although the Indonesian language is spoken by almost 200 million people and the 10th most spoken language in the world, it is under-represented in NLP research. Previous work on In…

cs.CL2020

WikiUMLS: Aligning UMLS to Wikipedia via Cross-lingual Neural Ranking

Afshin Rahimi, Timothy Baldwin, Karin Verspoor

We present our work on aligning the Unified Medical Language System (UMLS) to Wikipedia, to facilitate manual alignment of the two resources. We propose a cross-lingual neural rera…

cs.CL2019

Massively Multilingual Transfer for NER

Afshin Rahimi, Yuan Li, Trevor Cohn

In cross-lingual transfer, NLP models over one or more source languages are applied to a low-resource target language. While most prior work has used a single source model or a few…

cs.CL2018

Semi-supervised User Geolocation via Graph Convolutional Networks

Afshin Rahimi, Trevor Cohn, Timothy Baldwin

Social media user geolocation is vital to many applications such as event detection. In this paper, we propose GCN, a multiview geolocation model based on Graph Convolutional Netwo…