4 papers
Cross-Layer Misalignment Detection in Agent Skills: A Progressive Loading-Aware Contrastive Learning Approach
Chengjun Zhang, Yang Gao, Jianna Hur +2
Large language model (LLM) agents are increasingly extended through Agent Skills, reusable artifacts that package natural-language metadata, procedural instructions, and execution-…
Cybersecurity is the True Frontier for Generative AI Success or Failure
Edward Raff, Maor Ashkenazi, Sagar Samtani +2
Cybersecurity is a real-life test-bed for many machine learning problems at once, especially when considering modern strides in using Large Language Models (LLMs) to automate proce…
HackerSignal: A Large-Scale Multi-Source Dataset Linking Hacker Community Discourse to the CVE Vulnerability Lifecycle
Benjamin M. Ampel, Sagar Samtani
We introduce HackerSignal, a benchmark for temporal out-of-distribution cyber threat intelligence (CTI) and cross-source CVE linkage. HackerSignal aggregates 7.45 million exact-ded…
What Do Machine Learning Researchers Mean by "Reproducible"?
Edward Raff, Michel Benaroch, Sagar Samtani +1
The concern that Artificial Intelligence (AI) and Machine Learning (ML) are entering a "reproducibility crisis" has spurred significant research in the past few years. Yet with eac…