2 papers
cs.AI2026
Democratic Preference Alignment via Sortition-Weighted RLHF
Suvadip Sana, Jinzhou Wu, Martin T. Wells
Whose values should AI systems learn? Preference based alignment methods like RLHF derive their training signal from human raters, yet these rater pools are typically convenience s…
cs.GT2025
Quantitative Relaxations of Arrow's Axioms
Suvadip Sana, Daniel Brous, Martin T. Wells +1
In this paper we develop a novel approach to relaxing Arrow's axioms for voting rules, addressing a long-standing critique in social choice theory. Classical axioms (often styled a…