5 citations · 7 across the 10 of their papers we have counts for
3 papers · 1 filter
SeRA: Self-Reviewing and Alignment of Large Language Models using Implicit Reward Margins
Jongwoo Ko, Saket Dingliwal, Bhavana Ganesh +3
Direct alignment algorithms (DAAs), such as direct preference optimization (DPO), have become popular alternatives for Reinforcement Learning from Human Feedback (RLHF) due to thei…
Covariate Distribution Aware Meta-learning
Amrith Setlur, Saket Dingliwal, Barnabas Poczos
Meta-learning has proven to be successful for few-shot learning across the regression, classification, and reinforcement learning paradigms. Recent approaches have adopted Bayesian…
Finding Input Characterizations for Output Properties in ReLU Neural Networks
Saket Dingliwal, Divyansh Pareek, Jatin Arora
Deep Neural Networks (DNNs) have emerged as a powerful mechanism and are being increasingly deployed in real-world safety-critical domains. Despite the widespread success, their co…