3 papers
cs.LG2026
RewardUQ: A Unified Framework for Uncertainty-Aware Reward Models
Daniel Yang, Samuel Stante, Florian Redhardt +5
Reward models are central to aligning large language models (LLMs) with human preferences. Yet most approaches rely on pointwise reward estimates that overlook the epistemic uncert…
cs.CV2025
TrajFlow: Multi-modal Motion Prediction via Flow Matching
Qi Yan, Brian Zhang, Yutong Zhang +8
Efficient and accurate motion prediction is crucial for ensuring safety and informed decision-making in autonomous driving, particularly under dynamic real-world conditions that ne…
cs.LG2025
Near-optimal Active Reconstruction
Daniel Yang
With the growing practical interest in vision-based tasks for autonomous systems, the need for efficient and complex methods becomes increasingly larger. In the rush to develop new…