3 papers
cs.CL2024
AIPO: Improving Training Objective for Iterative Preference Optimization
Yaojie Shen, Xinyao Wang, Yulei Niu +5
Preference Optimization (PO), is gaining popularity as an alternative choice of Proximal Policy Optimization (PPO) for aligning Large Language Models (LLMs). Recent research on ali…
cs.IR2023
Steam Recommendation System
Samin Batra, Varun Sharma, Yurou Sun +2
We aim to leverage the interactions between users and items in the Steam community to build a game recommendation system that makes personalized suggestions to players in order to…
cs.IT2022
Secure UAV-to-Ground MIMO Communications: Joint Transceiver and Location Optimization
Zhong Zheng, Xinyao Wang, Zesong Fei +3
Unmanned aerial vehicles (UAVs) are foreseen to constitute promising airborne communication devices as a benefit of their superior channel quality. But UAV-to-ground (U2G) communic…