3 papers
cs.AI2026
Boosting In-Silicon Directed Evolution with Fine-Tuned Protein Language Model and Tree Search
Yaodong Yang, Yang Wang, Jinpeng Li +4
Protein evolution through amino acid mutations is a cornerstone of life sciences. Recent advances in protein language models have shown rich evolutionary patterns, offering unprece…
cs.MA2025
Dual Ensembled Multiagent Q-Learning with Hypernet Regularizer
Yaodong Yang, Guangyong Chen, Hongyao Tang +3
Overestimation in single-agent reinforcement learning has been extensively studied. In contrast, overestimation in the multiagent setting has received comparatively little attentio…
cs.LG2024
Safe Reinforcement Learning using Finite-Horizon Gradient-based Estimation
Juntao Dai, Yaodong Yang, Qian Zheng +1
A key aspect of Safe Reinforcement Learning (Safe RL) involves estimating the constraint condition for the next policy, which is crucial for guiding the optimization of safe policy…