Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
CritiQ: Mining Data Quality Criteria from Human Preferences
Honglin Guo, Kai Lv, Qipeng Guo +8
Language model heavily depends on high-quality data for optimal performance. Existing approaches rely on manually designed heuristics, the perplexity of existing models, training c…
cs.CL2024
F-Eval: Assessing Fundamental Abilities with Refined Evaluation Methods
Yu Sun, Keyu Chen, Shujie Wang +6
Large language models (LLMs) garner significant attention for their unprecedented performance, leading to an increasing number of researches evaluating LLMs. However, these evaluat…