4 citations · 5 across the 3 of their papers we have counts for
4 papers · 1 filter
Diffusion Model Alignment Using Direct Preference Optimization
Bram Wallace, Meihua Dang, Rafael Rafailov +7
Large language models (LLMs) are fine-tuned using human comparison data with Reinforcement Learning from Human Feedback (RLHF) methods to make them better aligned with users' prefe…
Activation Regression for Continuous Domain Generalization with Applications to Crop Classification
Samar Khanna, Bram Wallace, Kavita Bala +1
Geographic variance in satellite imagery impacts the ability of machine learning models to generalise to new regions. In this paper, we model geographic generalisation in medium re…
Extending and Analyzing Self-Supervised Learning Across Domains
Bram Wallace, Bharath Hariharan
Self-supervised representation learning has achieved impressive results in recent years, with experiments primarily coming on ImageNet or other similarly large internet imagery dat…
Few-Shot Generalization for Single-Image 3D Reconstruction via Priors
Bram Wallace, Bharath Hariharan
Recent work on single-view 3D reconstruction shows impressive results, but has been restricted to a few fixed categories where extensive training data is available. The problem of…