11 citations · 14 across the 13 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
Benchmarking LLMs' Mathematical Reasoning with Unseen Random Variables Questions
Zijin Hong, Hao Wu, Su Dong +8
Recent studies have raised significant concerns regarding the reliability of current mathematics benchmarks, highlighting issues such as simplistic design and potential data contam…
cs.CL2024
Unconstrained Model Merging for Enhanced LLM Reasoning
Yiming Zhang, Baoyi He, Shengyu Zhang +12
Recent advancements in building domain-specific large language models (LLMs) have shown remarkable success, especially in tasks requiring reasoning abilities like logical inference…
cs.CL2024
Collapsed Language Models Promote Fairness
Jingxuan Xu, Wuyang Chen, Linyi Li +2
To mitigate societal biases implicitly encoded in recent successful pretrained language models, a diverse array of approaches have been proposed to encourage model fairness, focusi…