activity
20232026
most citedLLaMAntino: LLaMA 2 Models for Effective Text Generation in Italian Language

10 citations · 10 across the 4 of their papers we have counts for

collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL2026

Mimir: Large-scale Multilingual Concept Modeling

Elio Musacchio, Lucia Siciliani, Pierpaolo Basile

Current language modeling approaches are built around tokens. Text corpora are split into tokens, and models are trained by performing computations on these tokens, such as predict…

cs.CL2025

Challenging the Abilities of Large Language Models in Italian: a Community Initiative

Malvina Nissim, Danilo Croce, Viviana Patti +78

The rapid progress of Large Language Models (LLMs) has transformed natural language processing and broadened its impact across research and society. Yet, systematic evaluation of t…

cs.CL2025

xVLM2Vec: Adapting LVLM-based embedding models to multilinguality using Self-Knowledge Distillation

Elio Musacchio, Lucia Siciliani, Pierpaolo Basile +1

In the current literature, most embedding models are based on the encoder-only transformer architecture to extract a dense and meaningful representation of the given input, which c…

cs.CL2025

Exploring the Word Sense Disambiguation Capabilities of Large Language Models

Pierpaolo Basile, Lucia Siciliani, Elio Musacchio +1

Word Sense Disambiguation (WSD) is a historical task in computational linguistics that has received much attention over the years. However, with the advent of Large Language Models…

cs.CL2023★ 10 cited

LLaMAntino: LLaMA 2 Models for Effective Text Generation in Italian Language

Pierpaolo Basile, Elio Musacchio, Marco Polignano +3

Large Language Models represent state-of-the-art linguistic models designed to equip computers with the ability to comprehend natural language. With its exceptional capacity to cap…