1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2025★ 1 cited
Aspect-Guided Multi-Level Perturbation Analysis of Large Language Models in Automated Peer Review
Jiatao Li, Yanheng Li, Xinyu Hu +2
We propose an aspect-guided, multi-level perturbation framework to evaluate the robustness of Large Language Models (LLMs) in automated peer review. Our framework explores perturba…
cs.CL2025
Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference
Mingqi Gao, Yixin Liu, Xinyu Hu +3
Evaluating and ranking the capabilities of different LLMs is crucial for understanding their performance and alignment with human preferences. Due to the high cost and time-consumi…