Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy
Peiyao Xiao, Xiaogang Li, Xinyi Gao +7
As LLMs achieved breakthroughs in general reasoning, their proficiency in specialized scientific domains reveals pronounced gaps in existing benchmarks due to data contamination, i…
cs.AI2026
CrystalXRD-Bench: Benchmarking Vision-Language Models for XRD Peak Indexing Across Diverse Crystalline Materials
Chengliang Xu, Xiaogang Li, Peiyao Xiao +3
Miller-index identification from powder XRD patterns requires capabilities untested by existing multimodal benchmarks: the model must read a narrow peak location from a rendered sc…