Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Can LLMs Really Understand Item Difficulty Levels? Implications for Automated Item Generation Using LLMs
Xinyi Wang, Hong Jiao, Ming Li +4
The estimation of item difficulty plays a key role in both formative assessment and large-scale high-stakes summative assessments. This study explores how large language models (LL…
cs.CL2024
Guiding Language Model Reasoning with Planning Tokens
Xinyi Wang, Lucas Caccia, Oleksiy Ostapenko +3
Large language models (LLMs) have recently attracted considerable interest for their ability to perform complex reasoning tasks, such as chain-of-thought (CoT) reasoning. However,…