2 papers
math.ST2026
Minimaxity and Admissibility of Bayesian Neural Networks
Daniel Andrew Coulson, Martin T. Wells
Bayesian neural networks (BNNs) offer a natural probabilistic formulation for inference in deep learning models. Despite their popularity, their optimality has received limited att…
cs.LG2025
Online Distributionally Robust LLM Alignment via Regression to Relative Reward
Sharan Sahu, Martin T. Wells
Reinforcement Learning with Human Feedback (RLHF) has become crucial for aligning Large Language Models (LLMs) with human intent. However, existing offline RLHF approaches suffer f…