2 papers
cs.CL2025
AURORA:Automated Training Framework of Universal Process Reward Models via Ensemble Prompting and Reverse Verification
Xiaoyu Tan, Tianchu Yao, Chao Qu +8
The reasoning capabilities of advanced large language models (LLMs) like o1 have revolutionized artificial intelligence applications. Nevertheless, evaluating and optimizing comple…
cs.LG2024
Robust Deep Hawkes Process under Label Noise of Both Event and Occurrence
Xiaoyu Tan, Bin Li, Xihe Qiu +3
Integrating deep neural networks with the Hawkes process has significantly improved predictive capabilities in finance, health informatics, and information technology. Nevertheless…