5 papers
Light or Full Verb? A Minimal-Pair Dataset for Probing Phraseological Competence in Language Models
Francesca Franzon, Nicolas Rosàs Gómez, Leo Wanner
Frequent English verbs such as 'have' and 'make' can function either as collocates in light-verb constructions or as full lexical predicates, as in 'make a decision' vs. 'make a ca…
Tracing the complexity profiles of different linguistic phenomena through the intrinsic dimension of LLM representations
Marco Baroni, Emily Cheng, Iria de-Dios-Flores +1
We explore intrinsic dimension (ID) of LLM representations as a marker of linguistic complexity. Specifically, we test whether ID differences across model layers reflect well-known…
The emergence of numerical representations in communicating artificial agents
Daniela Mihai, Lucas Weber, Francesca Franzon
Human languages provide efficient systems for expressing numerosities, but whether the sheer pressure to communicate is enough for numerical representations to arise in artificial…
Repetitions are not all alike: distinct mechanisms sustain repetition in language models
Matéo Mahaut, Francesca Franzon
Large Language Models (LLMs) can sometimes degrade into repetitive loops, persistently generating identical word sequences. Because repetition is rare in natural human language, it…
Principles of semantic and functional efficiency in grammatical patterning
Emily Cheng, Francesca Franzon
Grammatical features such as number and gender serve two central functions in human languages. While they encode salient semantic attributes like numerosity and animacy, they also…