1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CL2024
RS-DPO: A Hybrid Rejection Sampling and Direct Preference Optimization Method for Alignment of Large Language Models
Saeed Khaki, JinJin Li, Lan Ma +2
Reinforcement learning from human feedback (RLHF) has been extensively employed to align large language models with user intent. However, proximal policy optimization (PPO) based R…
cs.SD2024
Exploring Federated Self-Supervised Learning for General Purpose Audio Understanding
Yasar Abbas Ur Rehman, Kin Wai Lau, Yuyang Xie +2
The integration of Federated Learning (FL) and Self-supervised Learning (SSL) offers a unique and synergetic combination to exploit the audio data for general-purpose audio underst…
cs.CV2023★ 1 cited
Texture Generation on 3D Meshes with Point-UV Diffusion
Xin Yu, Peng Dai, Wenbo Li +3
In this work, we focus on synthesizing high-quality textures on 3D meshes. We present Point-UV diffusion, a coarse-to-fine pipeline that marries the denoising diffusion model with…