2 papers
cs.SE2025
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
Shuo Yan, Ruochen Li, Ziming Luo +11
Large language model (LLM) agents have demonstrated remarkable potential in advancing scientific discovery. However, their capability in the fundamental yet crucial task of reprodu…
cs.CL2025
Exploring Multilingual Probing in Large Language Models: A Cross-Language Analysis
Daoyang Li, Haiyan Zhao, Qingcheng Zeng +1
Probing techniques for large language models (LLMs) have primarily focused on English, overlooking the vast majority of the world's languages. In this paper, we extend these probin…