2 papers
cs.LG2025
Trust Region Reward Optimization and Proximal Inverse Reward Optimization Algorithm
Yang Chen, Menglin Zou, Jiaqi Zhang +6
Inverse Reinforcement Learning (IRL) learns a reward function to explain expert demonstrations. Modern IRL methods often use the adversarial (minimax) formulation that alternates b…
cs.CE2025
Multi-source Multi-level Multi-token Ethereum Dataset and Benchmark Platform
Haoyuan Li, Mengxiao Zhang, Maoyuan Li +5
This paper introduces 3MEthTaskforce (https://3meth.github.io), a multi-source, multi-level, and multi-token Ethereum dataset addressing the limitations of single-source datasets.…