Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Not All Preference Pairs Are Created Equal: A Recipe for Annotation-Efficient Iterative Preference Learning
Sen Yang, Leyang Cui, Deng Cai +3
Iterative preference learning, though yielding superior performances, requires online annotated preference labels. In this work, we study strategies to select worth-annotating resp…
cs.CL2023
Once Upon a in : Relative-Time Pretraining for Complex Temporal Reasoning
Sen Yang, Xin Li, Lidong Bing +1
Our physical world is constantly evolving over time, rendering challenges for pre-trained language models to understand and reason over the temporal contexts of texts. Existing wor…