2 papers
cs.CL2026
Fixing the Broken Compass: Diagnosing and Improving Inference-Time Reward Modeling
Jiachun Li, Pengfei Cao, Zhuoran Jin +6
Inference-time scaling techniques have shown promise in enhancing the reasoning capabilities of large language models (LLMs). While recent research has primarily focused on trainin…
cs.CL2024
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
Jiachun Li, Pengfei Cao, Chenhao Wang +6
Large language models (LLMs) sometimes demonstrate poor performance on knowledge-intensive tasks, commonsense reasoning is one of them. Researchers typically address these issues b…