Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
AgriEval: A Comprehensive Chinese Agricultural Benchmark for Large Language Models
Lian Yan, Haotian Wang, Chen Tang +5
In the agricultural domain, the deployment of large language models (LLMs) is hindered by the lack of training data and evaluation benchmarks. To mitigate this issue, we propose Ag…
cs.CL2024
Learning to Break: Knowledge-Enhanced Reasoning in Multi-Agent Debate System
Haotian Wang, Xiyuan Du, Weijiang Yu +5
Multi-agent debate system (MAD) imitating the process of human discussion in pursuit of truth, aims to align the correct cognition of different agents for the optimal solution. It…