2 papers
cs.CL2025
Your Finetuned Large Language Model is Already a Powerful Out-of-distribution Detector
Andi Zhang, Tim Z. Xiao, Weiyang Liu +2
We revisit the likelihood ratio between a pretrained large language model (LLM) and its finetuned variant as a criterion for out-of-distribution (OOD) detection. The intuition behi…
stat.ML2024
Constructing Semantics-Aware Adversarial Examples with a Probabilistic Perspective
Andi Zhang, Mingtian Zhang, Damon Wischik
We propose a probabilistic perspective on adversarial examples, allowing us to embed subjective understanding of semantics as a distribution into the process of generating adversar…