20 citations · 20 across the 3 of their papers we have counts for
3 papers
cs.CL2023
Tokenization with Factorized Subword Encoding
David Samuel, Lilja Øvrelid
In recent years, language models have become increasingly larger and more complex. However, the input representations for these models continue to rely on simple and greedy subword…
cs.CL2023
Text-To-KG Alignment: Comparing Current Methods on Classification Tasks
Sondre Wold, Lilja Øvrelid, Erik Velldal
In contrast to large text corpora, knowledge graphs (KG) provide dense and structured representations of factual information. This makes them attractive for systems that supplement…
cs.CL2022★ 20 cited
Contextualized language models for semantic change detection: lessons learned
Andrey Kutuzov, Erik Velldal, Lilja Øvrelid
We present a qualitative analysis of the (potentially erroneous) outputs of contextualized embedding-based methods for detecting diachronic semantic change. First, we introduce an…