271 citations · 387 across the 28 of their papers we have counts for
6 papers · 1 filter
TACIT: A Target-Agnostic Feature Disentanglement Framework for Cross-Domain Text Classification
Rui Song, Fausto Giunchiglia, Yingji Li +2
Cross-domain text classification aims to transfer models from label-rich source domains to label-poor target domains, giving it a wide range of practical applications. Many approac…
Lexical Diversity in Kinship Across Languages and Dialects
Hadi Khalilia, Gábor Bella, Abed Alhakim Freihat +2
Languages are known to describe the world in diverse ways. Across lexicons, diversity is pervasive, appearing through phenomena such as lexical gaps and untranslatability. However,…
Towards Bridging the Digital Language Divide
Gábor Bella, Paula Helm, Gertraud Koch +1
It is a well-known fact that current AI-based language technology -- language models, machine translation systems, multilingual dictionaries and corpora -- focuses on the world's 2…
Automatic Counterfactual Augmentation for Robust Text Classification Based on Word-Group Search
Rui Song, Fausto Giunchiglia, Yingji Li +1
Despite large-scale pre-trained language models have achieved striking results for text classificaion, recent work has raised concerns about the challenge of shortcut learning. In…
Using Linguistic Typology to Enrich Multilingual Lexicons: the Case of Lexical Gaps in Kinship
Temuulen Khishigsuren, Gábor Bella, Khuyagbaatar Batsuren +6
This paper describes a method to enrich lexical resources with content relating to linguistic diversity, based on knowledge from the field of lexical typology. We capture the pheno…
Language Diversity: Visible to Humans, Exploitable by Machines
Gábor Bella, Erdenebileg Byambadorj, Yamini Chandrashekar +3
The Universal Knowledge Core (UKC) is a large multilingual lexical database with a focus on language diversity and covering over a thousand languages. The aim of the database, as w…