4 papers
Streaming Generated Gaussian Process Experts for Online Learning and Control: Extended Version
Zewen Yang, Dongfa Zhang, Xiaobing Dai +5
Gaussian Processes (GPs), as a nonparametric learning method, offer flexible modeling capabilities and calibrated uncertainty quantification for function approximations. Additional…
Pessimism Principle Can Be Effective: Towards a Framework for Zero-Shot Transfer Reinforcement Learning
Chi Zhang, Ziying Jia, George K. Atia +2
Transfer reinforcement learning aims to derive a near-optimal policy for a target environment with limited data by leveraging abundant data from related source domains. However, it…
DualOpt: A Dual Divide-and-Optimize Algorithm for the Large-scale Traveling Salesman Problem
Shipei Zhou, Yuandong Ding, Chi Zhang +2
This paper proposes a dual divide-and-optimize algorithm (DualOpt) for solving the large-scale traveling salesman problem (TSP). DualOpt combines two complementary strategies to im…
Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline
Guancheng Zeng, Wentao Ding, Beining Xu +8
Enterprises possess a vast array of API assets scattered across various functions, forming the backbone of existing business processes. By leveraging these APIs as functional tools…