1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CR2024
Kov: Transferable and Naturalistic Black-Box LLM Attacks using Markov Decision Processes and Tree Search
Robert J. Moss
Eliciting harmful behavior from large language models (LLMs) is an important task to ensure the proper alignment and safety of the models. Often when training LLMs, ethical guideli…
cs.LG2023★ 1 cited
Formal and Practical Elements for the Certification of Machine Learning Systems
Jean-Guillaume Durand, Arthur Dubois, Robert J. Moss
Over the past decade, machine learning has demonstrated impressive results, often surpassing human capabilities in sensing tasks relevant to autonomous flight. Unlike traditional a…