1 citations · 1 across the 4 of their papers we have counts for
4 papers
-Puzzle: A Cost-Efficient Testbed for Benchmarking Reinforcement Learning Algorithms in Generative Language Model
Yufeng Zhang, Liyu Chen, Boyi Liu +4
Recent advances in reinforcement learning (RL) algorithms aim to enhance the performance of language models at scale. Yet, there is a noticeable absence of a cost-effective and sta…
Differentiable Arbitrating in Zero-sum Markov Games
Jing Wang, Meichen Song, Feng Gao +3
We initiate the study of how to perturb the reward in a zero-sum Markov game with two players to induce a desirable Nash equilibrium, namely arbitrating. Such a problem admits a bi…
An Efficient Approach to the Online Multi-Agent Path Finding Problem by Using Sustainable Information
Mingkai Tang, Boyi Liu, Yuanhang Li +3
Multi-agent path finding (MAPF) is the problem of moving agents to the goal vertex without collision. In the online MAPF problem, new agents may be added to the environment at any…
AuthROS: Secure Data Sharing Among Robot Operating Systems Based on Ethereum
Shenhui Zhang, Wenkai Li, Xiaoqi Li +1
The Robot Operating System (ROS) streamlines human processes, increasing the efficiency of various production tasks. However, the security of data transfer operations in ROS is sti…