3 citations · 3 across the 1 of their papers we have counts for
3 papers
Cross-lingual Human-Preference Alignment for Neural Machine Translation with Direct Quality Optimization
Kaden Uhlig, Joern Wuebker, Raphael Reinauer +1
Reinforcement Learning from Human Feedback (RLHF) and derivative techniques like Direct Preference Optimization (DPO) are task-alignment algorithms used to repurpose general, found…
Neural Machine Translation Models Can Learn to be Few-shot Learners
Raphael Reinauer, Patrick Simianer, Kaden Uhlig +2
The emergent ability of Large Language Models to use a small number of examples to learn to perform in novel domains and tasks, also called in-context learning (ICL). In this work,…
Early Warning Signals of Social Instabilities in Twitter Data
Vahid Shamsaddini, Henry Kirveslahti, Raphael Reinauer +3
The goal of this project is to create and study novel techniques to identify early warning signals for socially disruptive events, like riots, wars, or revolutions using only publi…