2 papers
cs.CR2025
Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risks of Language Models
Andy K. Zhang, Neil Perry, Riya Dulepet +24
Language Model (LM) agents for cybersecurity that are capable of autonomously identifying vulnerabilities and executing exploits have potential to cause real-world impact. Policyma…
cs.LG2025
Preferential Multi-Objective Bayesian Optimization for Drug Discovery
Tai Dang, Long-Hung Pham, Sang T. Truong +6
Despite decades of advancements in automated ligand screening, large-scale drug discovery remains resource-intensive and requires post-processing hit selection, a step where chemis…