50 citations · 52 across the 6 of their papers we have counts for
4 papers · 1 filter
Adaptive Nucleus Truncation for Long-Form Reasoning
Ousmane Amadou Dia
Sampling plays an important role in long-form language-model reasoning. Over thousands of decoding steps, small changes in the candidate token set can compound into different reaso…
Variational Proximal Policy Optimization
Ousmane Amadou Dia
Reinforcement Learning from Human Feedback via Proximal Policy Optimization often suffers from policy mode collapse, brittle exploration loops, and distribution drift. This paper i…
Localized Uncertainty Attacks
Ousmane Amadou Dia, Theofanis Karaletsos, Caner Hazirbas +3
The susceptibility of deep learning models to adversarial perturbations has stirred renewed attention in adversarial examples resulting in a number of attacks. However, most of the…
Semantics Preserving Adversarial Learning
Ousmane Amadou Dia, Elnaz Barshan, Reza Babanezhad
While progress has been made in crafting visually imperceptible adversarial examples, constructing semantically meaningful ones remains a challenge. In this paper, we propose a fra…