2 papers
cs.LG2026
ProSpec RL: Plan Ahead, then Execute
Liangliang Liu, Yi Guan, BoRan Wang +5
Imagining potential outcomes of actions before execution helps agents make more informed decisions, a prospective thinking ability fundamental to human cognition. However, mainstre…
cs.CL2025
AgriEval: A Comprehensive Chinese Agricultural Benchmark for Large Language Models
Lian Yan, Haotian Wang, Chen Tang +5
In the agricultural domain, the deployment of large language models (LLMs) is hindered by the lack of training data and evaluation benchmarks. To mitigate this issue, we propose Ag…