2 papers
cs.CR2025
Never compromise with vulnerabilities: a comprehensive survey on AI governance
Yuchu Jiang, Jian Zhao, Yuchen Yuan +64
The rapid advancement of AI has expanded its capabilities across domains, yet introduced critical technical vulnerabilities, such as algorithmic bias and adversarial sensitivity, t…
cs.CL2025
Aligning Black-box Language Models with Human Judgments
Gerrit J. J. van den Burg, Gen Suzuki, Wei Liu +1
Large language models (LLMs) are increasingly used as automated judges to evaluate recommendation systems, search engines, and other subjective tasks, where relying on human evalua…