10 citations · 10 across the 2 of their papers we have counts for
2 papers
cs.AI2026
Agentic BAIM-LLM Evaluation (ABLE): Benchmarking LLM Use of Protein Design Tools
Bryce Cai, Geetha Jeyapragasan, Samira Nedungadi +2
We introduce ABLE, a benchmark for evaluating LLM agents' ability to use biological AI models (BAIMs), such as ProteinMPNN and AlphaFold3, in dual-use protein design workflows. ABL…
cs.AI2023★ 10 cited
Will releasing the weights of future large language models grant widespread access to pandemic agents?
Anjali Gopal, Nathan Helm-Burger, Lennart Justen +6
Large language models can benefit research and human understanding by providing tutorials that draw on expertise from many different fields. A properly safeguarded model will refus…