3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.LG2024
Sample-Efficient Alignment for LLMs
Zichen Liu, Changyu Chen, Chao Du +2
We study methods for efficiently aligning large language models (LLMs) with human preferences given budgeted online feedback. We first formulate the LLM alignment problem in the fr…
quant-ph2024★ 1 cited
Long-distance distribution of telecom time-energy entanglement generated on a silicon chip
Yuan-yuan Zhao, Fuyong Yue, Feng Gao +5
Entanglement distribution is a critical technique that enables numerous quantum applications. Most fiber-based long-distance experiments reported to date have utilized photon pair…
cs.LG2024★ 3 cited
Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning
Shengyi Huang, Quentin Gallouédec, Florian Felten +30
In many Reinforcement Learning (RL) papers, learning curves are useful indicators to measure the effectiveness of RL algorithms. However, the complete raw data of the learning curv…