benchmarking 1conflict resolution 1constraint compliance 1instruction hierarchy 1llm robustness 1tool use 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CR2026
IH-Benchmark: A Conflict-Centered Benchmark for Instruction-Hierarchy Robustness in LLM Applications
Conor McCauley, Zeliang Kan, Jason Martin
The paper introduces IH-Benchmark, a dataset for evaluating how large language models handle conflicting instructions across system‑user and user‑tool hierarchies, using a taxonomy…
cs.CR2026
Beyond the TESSERACT:Trustworthy Dataset Curation for Sound Evaluations of Android Malware Classifiers
Theo Chow, Mario D'Onghia, Lorenz Linhardt +4
The reliability of machine learning critically depends on dataset quality. While machine learning applied to computer vision and natural language processing benefits from high-qual…
cs.LG2025
TESSERACT: Eliminating Experimental Bias in Malware Classification across Space and Time (Extended Version)
Zeliang Kan, Shae McFadden, Daniel Arp +5
Machine learning (ML) plays a pivotal role in detecting malicious software. Despite the high F1-scores reported in numerous studies reaching upwards of 0.99, the issue is not compl…