Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
DefenderBench: A Toolkit for Evaluating Language Agents in Cybersecurity Environments
Chiyu Zhang, Marc-Alexandre Cote, Michael Albada +6
Large language model (LLM) agents have shown impressive capabilities in human language comprehension and reasoning, yet their potential in cybersecurity remains underexplored. We i…
cs.CL2025
Group Preference Alignment: Customized LLM Response Generation from In-Situ Conversations
Ishani Mondal, Jack W. Stokes, Sujay Kumar Jauhar +5
LLMs often fail to meet the specialized needs of distinct user groups due to their one-size-fits-all training paradigm \cite{lucy-etal-2024-one} and there is limited research on wh…