2 papers
cs.LG2026
Uncertainty-Aware Trust Estimation for Multi-LLM Systems via Structured Expert Judgement
Jiawei Zheng, Jiazhen Zhang
Large Language Model (LLM) ensembles are increasingly used to improve reliability by combining predictions from multiple LLMs. However, existing aggregation methods typically assum…
cs.SE2026
Large Language Models for Multi-Lingual Equivalent Mutant Detection: An Extended Empirical Study
Honglin Shu, Zhao Tian, Dong Wang +5
Mutation testing is a powerful technique for ensuring software quality. However, the presence of equivalent mutants introduces unnecessary costs and biases, limiting its practical…