2 papers
cs.LG2026
Represent, Then Generate: Multimodal-Conditioned Time-Series Generation under Irregular Missingness
Haochen Zhang, Jiaheng Guo, Yu-Chao Huang +3
Continuous physiological time series underpin modern clinical monitoring, yet many of the most informative signals are invasive, expensive, or simply unavailable for a given patien…
cs.CL2024
Improving Reward Models with Synthetic Critiques
Zihuiwen Ye, Fraser Greenlee-Scott, Max Bartolo +3
Reward models (RMs) play a critical role in aligning language models through the process of reinforcement learning from human feedback. RMs are trained to predict a score reflectin…