gradient-free methods 1in-context learning 1multimodal retrieval 1on-device inference 1task-conditioned retrieval 1
From the 1 of 17 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Unbiased Alignment for Large Language Models with Noisy Preferences
Jialiang Wang, Xianming Liu, Xiong Zhou +2
The alignment of large language models with human preferences is commonly achieved through Reinforcement Learning from Human Feedback or Direct Preference Optimization. However, th…
cs.LG2024
Unraveling the Mechanics of Learning-Based Demonstration Selection for In-Context Learning
Hui Liu, Wenya Wang, Hao Sun +4
Large Language Models (LLMs) have demonstrated impressive in-context learning (ICL) capabilities from few-shot demonstration exemplars. While recent learning-based demonstration se…