Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
PRD: Peer Rank and Discussion Improve Large Language Model based Evaluations
Ruosen Li, Teerth Patel, Xinya Du
Nowadays, the quality of responses generated by different modern large language models (LLMs) is hard to evaluate and compare automatically. Recent studies suggest and predominantl…
cs.CL2024
Large Language Models for Automated Open-domain Scientific Hypotheses Discovery
Zonglin Yang, Xinya Du, Junxian Li +3
Hypothetical induction is recognized as the main reasoning type when scientists make observations about the world and try to propose hypotheses to explain those observations. Past…