3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2025
Rethinking LLM Evaluation: Can We Evaluate LLMs with 200x Less Data?
Shaobo Wang, Cong Wang, Wenjie Fu +11
As the demand for comprehensive evaluations of diverse model capabilities steadily increases, benchmark suites have correspondingly grown significantly in scale. Despite notable ad…
cs.AI2025★ 3 cited
A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy
Huandong Wang, Wenjie Fu, Yingzhou Tang +7
While large language models (LLMs) present significant potential for supporting numerous real-world applications and delivering positive social impacts, they still face significant…