3 papers
cs.SE2025
AutoICE: Automatically Synthesizing Verifiable C Code via LLM-driven Evolution
Weilin Luo, Xueyi Liang, Haotian Deng +2
Automatically synthesizing verifiable code from natural language requirements ensures software correctness and reliability while significantly lowering the barrier to adopting the…
cs.CL2025
SCoRE: Benchmarking Long-Chain Reasoning in Commonsense Scenarios
Weidong Zhan, Yue Wang, Nan Hu +12
Currently, long-chain reasoning remains a key challenge for large language models (LLMs) because natural texts lack sufficient explicit reasoning data. However, existing benchmarks…
cs.AI2025
Reinforcement Learning with Knowledge Representation and Reasoning: A Brief Survey
Chao Yu, Shicheng Ye, Hankz Hankui Zhuo
Reinforcement Learning (RL) has achieved tremendous development in recent years, but still faces significant obstacles in addressing complex real-life problems due to the issues of…