1 citations · 1 across the 4 of their papers we have counts for
4 papers
TDRM: Smooth Reward Models with Temporal Difference for LLM RL and Inference
Dan Zhang, Min Cai, Jonathan Light +3
Reward models are central to both reinforcement learning (RL) with language models and inference-time verification. However, existing reward models often lack temporal consistency,…
Measurement of the space-like transition form factor
BESIII Collaboration, M. Ablikim, M. N. Achasov +689
Based on of collision data taken with the BESIII detector at a center-of-mass energy of , the two-photon fusion process $e^+e^-\t…
In-the-wild Audio Spatialization with Flexible Text-guided Localization
Tianrui Pan, Jie Liu, Zewen Huang +2
To enhance immersive experiences, binaural audio offers spatial awareness of sounding objects in AR, VR, and embodied AI applications. While existing audio spatialization methods c…
Brief Industry Paper: The Necessity of Adaptive Data Fusion in Infrastructure-Augmented Autonomous Driving System
Shaoshan Liu, Jianda Wang, Zhendong Wang +7
This paper is the first to provide a thorough system design overview along with the fusion methods selection criteria of a real-world cooperative autonomous driving system, named I…