Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Understanding Generalization through Decision Pattern Shift
Huiqi Deng, Yibo Li, Quanshi Zhang +3
Understanding why deep neural networks (DNNs) fail to generalize to unseen samples remains a long-standing challenge. Existing studies mainly examine changes in externally observab…
cs.LG2026
SafeSci: Safety Evaluation of Large Language Models in Science Domains and Beyond
Xiangyang Zhu, Yuan Tian, Qi Jia +14
The success of large language models (LLMs) in scientific domains has heightened safety concerns, prompting numerous benchmarks to evaluate their scientific safety. Existing benchm…