Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Uncertainty-Aware Reward Modeling for Stable RLHF
Licheng Pan, Haocheng Yang, Haoxuan Li +7
Reinforcement learning from human feedback (RLHF) aligns large language models by training reward models on preference data and optimizing policies to maximize predicted rewards. H…
cs.LG2024
Sora Detector: A Unified Hallucination Detection for Large Text-to-Video Models
Zhixuan Chu, Lei Zhang, Yichen Sun +4
The rapid advancement in text-to-video (T2V) generative models has enabled the synthesis of high-fidelity video content guided by textual descriptions. Despite this significant pro…