3 papers
cs.CL2025
Deep Literature Survey Automation with an Iterative Workflow
Hongbo Zhang, Han Cui, Yidong Wang +6
Automatic literature survey generation has attracted increasing attention, yet most existing systems follow a one-shot paradigm, where a large set of papers is retrieved at once an…
cs.AI2025
TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them
Yidong Wang, Yunze Song, Tingyuan Zhu +11
The adoption of Large Language Models (LLMs) as automated evaluators (LLM-as-a-judge) has revealed critical inconsistencies in current evaluation frameworks. We identify two fundam…
cs.CL2025
Dynamics of Instruction Fine-Tuning for Chinese Large Language Models
Chiyu Song, Zhanchao Zhou, Jianhao Yan +3
Instruction tuning is a burgeoning method to elicit the general intelligence of Large Language Models (LLMs). While numerous studies have examined the impact of factors such as dat…