2 papers
cs.LG2026
Benchmark Shadows: Data Alignment, Parameter Footprints, and Generalization in Large Language Models
Hongjian Zou, Yidan Wang, Qi Ding +2
Large language models often achieve strong benchmark gains without corresponding improvements in broader capability. We hypothesize that this discrepancy arises from differences in…
cs.CL2024
Instruct Large Language Models to Generate Scientific Literature Survey Step by Step
Yuxuan Lai, Yupeng Wu, Yidan Wang +2
Abstract. Automatically generating scientific literature surveys is a valuable task that can significantly enhance research efficiency. However, the diverse and complex nature of i…