3 papers
cs.CL2026
BenCSSmark: Making the Social Sciences Count in LLM Research
Arnault Chatelain, Ãtienne Ollion, Qianwen Guan +7
This position paper argues that the under-representation of social science tasks in contemporary LLM benchmarks limits advances in both LLM evaluation and social scientific inquiry…
cs.CL2026
Pantagruel: Unified Self-Supervised Encoders for French Text and Speech
Phuong-Hang Le, Valentin Pelloin, Arnault Chatelain +27
We release Pantagruel models, a new family of self-supervised encoder models for French text and speech. Instead of predicting modality-tailored targets such as textual tokens or s…
cs.CL2025
GETALP@AutoMin 2025: Leveraging RAG to Answer Questions based on Meeting Transcripts
Jeongwoo Kang, Markarit Vartampetian, Felix Herron +3
This paper documents GETALP's submission to the Third Run of the Automatic Minuting Shared Task at SIGDial 2025. We participated in Task B: question-answering based on meeting tran…