1 citations · 1 across the 4 of their papers we have counts for
4 papers · 1 filter
Clear Preferences Leave Traces: Reference Model-Guided Sampling for Preference Learning
Nirav Diwan, Tolga Ergen, Dongsub Shim +1
Direct Preference Optimization (DPO) has emerged as a de-facto approach for aligning language models with human preferences. Recent work has shown DPO's effectiveness relies on tra…
MultiPrompter: Cooperative Prompt Optimization with Multi-Agent Reinforcement Learning
Dong-Ki Kim, Sungryull Sohn, Lajanugen Logeswaran +2
Recently, there has been an increasing interest in automated prompt optimization based on reinforcement learning (RL). This approach offers important advantages, such as generating…
EDDA: Explanation-driven Data Augmentation to Improve Explanation Faithfulness
Ruiwen Li, Zhibo Zhang, Jiani Li +5
Recent years have seen the introduction of a range of methods for post-hoc explainability of image classifier predictions. However, these post-hoc explanations may not always be fa…
Online Class-Incremental Continual Learning with Adversarial Shapley Value
Dongsub Shim, Zheda Mai, Jihwan Jeong +3
As image-based deep learning becomes pervasive on every device, from cell phones to smart watches, there is a growing need to develop methods that continually learn from data while…