2 papers
cs.CL2026
BenCSSmark: Making the Social Sciences Count in LLM Research
Arnault Chatelain, Ãtienne Ollion, Qianwen Guan +7
This position paper argues that the under-representation of social science tasks in contemporary LLM benchmarks limits advances in both LLM evaluation and social scientific inquiry…
cs.CL2026
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
Phuong-Hang Le, Valentin Pelloin, Arnault Chatelain +27
We release Pantagruel models, a new family of self-supervised encoder models for French text and speech. Instead of predicting modality-tailored targets such as textual tokens or s…