2 papers
cs.CL2026
ChiEngMixBench: Evaluating Large Language Models on Expert-Style Chinese-English Terminology Mixing
Qingyan Yang, Tongxi Wang, Yunsheng Luo
Large language models increasingly mediate multilingual professional communication, where useful generation requires adapting to community conventions about which expressions are r…
cs.CV2025
Medical Imaging AI Competitions Lack Fairness
Annika Reinke, Evangelia Christodoulou, Sthuthi Sadananda +34
Benchmarking competitions are central to the development of artificial intelligence (AI) in medical imaging, defining performance standards and shaping methodological progress. How…