1 citations · 1 across the 18 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
STAR : Bridging Statistical and Agentic Reasoning for Large Model Performance Prediction
Xiaoxiao Wang, Chunxiao Li, Junying Wang +6
As comprehensive large model evaluation becomes prohibitively expensive, predicting model performance from limited observations has become essential. However, existing statistical…
cs.AI2025
Who is a Better Player: LLM against LLM
Yingjie Zhou, Jiezhang Cao, Farong Wen +10
Adversarial board games, as a paradigmatic domain of strategic reasoning and intelligence, have long served as both a popular competitive activity and a benchmark for evaluating ar…