1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
BPO: Towards Balanced Preference Optimization between Knowledge Breadth and Depth in Alignment
Sizhe Wang, Yongqi Tong, Hengyuan Zhang +3
Reinforcement Learning with Human Feedback (RLHF) is the key to the success of large language models (LLMs) in recent years. In this work, we first introduce the concepts of knowle…
cs.LG2024
FlowTS: Time Series Generation via Rectified Flow
Yang Hu, Xiao Wang, Zezhen Ding +7
Diffusion-based models have significant achievements in time series generation but suffer from inefficient computation: solving high-dimensional ODEs/SDEs via iterative numerical s…