4 papers
Mimir: Large-scale Multilingual Concept Modeling
Elio Musacchio, Lucia Siciliani, Pierpaolo Basile
Current language modeling approaches are built around tokens. Text corpora are split into tokens, and models are trained by performing computations on these tokens, such as predict…
Challenging the Abilities of Large Language Models in Italian: a Community Initiative
Malvina Nissim, Danilo Croce, Viviana Patti +78
The rapid progress of Large Language Models (LLMs) has transformed natural language processing and broadened its impact across research and society. Yet, systematic evaluation of t…
xVLM2Vec: Adapting LVLM-based embedding models to multilinguality using Self-Knowledge Distillation
Elio Musacchio, Lucia Siciliani, Pierpaolo Basile +1
In the current literature, most embedding models are based on the encoder-only transformer architecture to extract a dense and meaningful representation of the given input, which c…
Exploring the Word Sense Disambiguation Capabilities of Large Language Models
Pierpaolo Basile, Lucia Siciliani, Elio Musacchio +1
Word Sense Disambiguation (WSD) is a historical task in computational linguistics that has received much attention over the years. However, with the advent of Large Language Models…