2 papers
stat.ME2026
Estimating Item Difficulty with Large Language Models as Experts
Diana Kolesnikova, Kirill Fedyanin, Abe D. Hofman +2
Accurate estimates of item difficulty are essential for valid assessment and effective adaptive learning. However, for newly created tasks, response data are typically unavailable.…
cs.CL2024
Falcon2-11B Technical Report
Quentin Malartic, Nilabhra Roy Chowdhury, Ruxandra Cojocaru +14
We introduce Falcon2-11B, a foundation model trained on over five trillion tokens, and its multimodal counterpart, Falcon2-11B-vlm, which is a vision-to-text model. We report our f…