2 papers
cs.CL2026
RankLLM: Weighted Ranking of LLMs by Quantifying Question Difficulty
Ziqian Zhang, Xingjian Hu, Yue Huang +8
Benchmarks establish a standardized evaluation framework to systematically assess the performance of large language models (LLMs), facilitating objective comparisons and driving ad…
cond-mat.mtrl-sci2025
Zero-shot Autonomous Microscopy for Scalable and Intelligent Characterization of 2D Materials
Jingyun Yang, Ruoyan Avery Yin, Chi Jiang +14
Characterization of atomic-scale materials traditionally requires human experts with months to years of specialized training. Even for trained human operators, accurate and reliabl…