Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024
Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Usman Anwar, Abulhair Saparov, Javier Rando +39
This work identifies 18 foundational challenges in assuring the alignment and safety of large language models (LLMs). These challenges are organized into three different categories…
cs.LG2021
Neural Variational Gradient Descent
Lauro Langosco di Langosco, Vincent Fortuin, Heiko Strathmann
Particle-based approximate Bayesian inference approaches such as Stein Variational Gradient Descent (SVGD) combine the flexibility and convergence guarantees of sampling methods wi…