3 papers
cs.CR2026
IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications
Conor McCauley, Zeliang Kan, Jason Martin
When a language model receives conflicting instructions from different priority levels, which one does it actually follow? This question lies at the heart of reliable LLM deploymen…
cs.CR2025
Beyond the TESSERACT:Trustworthy Dataset Curation for Sound Evaluations of Android Malware Classifiers
Theo Chow, Mario D'Onghia, Lorenz Linhardt +4
The reliability of machine learning critically depends on dataset quality. While machine learning applied to computer vision and natural language processing benefits from high-qual…
cs.CR2019
Automated Deobfuscation of Android Native Binary Code
Zeliang Kan, Haoyu Wang, Lei Wu +2
With the popularity of Android apps, different techniques have been proposed to enhance app protection. As an effective approach to prevent reverse engineering, obfuscation can be…