Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Toward Understanding In-context vs. In-weight Learning
Bryan Chan, Xinyi Chen, András György +1
It has recently been demonstrated empirically that in-context learning emerges in transformers when certain distributional properties are present in the training data, but this abi…
cs.LG2024
Preference Learning Algorithms Do Not Learn Preference Rankings
Angelica Chen, Sadhika Malladi, Lily H. Zhang +4
Preference learning algorithms (e.g., RLHF and DPO) are frequently used to steer LLMs to produce generations that are more preferred by humans, but our understanding of their inner…