Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
SpidR-Adapt: A Universal Speech Representation Model for Few-Shot Adaptation
Mahi Luthra, Jiayi Shen, Maxime Poli +14
Human infants, with only a few hundred hours of speech exposure, acquire basic units of new languages, highlighting a striking efficiency gap compared to the data-hungry self-super…
cs.CL2025
LongTail-Swap: benchmarking language models' abilities on rare words
Robin Algayres, Charles-Ãric Saint-James, Mahi Luthra +6
Children learn to speak with a low amount of data and can be taught new words on a few-shot basis, making them particularly data-efficient learners. The BabyLM challenge aims at ex…