3 papers
cs.LG2026
Active Advantage-Aligned Online Reinforcement Learning with Offline Data
Xuefeng Liu, Hung T. C. Le, Siyu Chen +4
Online reinforcement learning (RL) enhances policies through direct interactions with the environment, but faces challenges related to sample efficiency. In contrast, offline RL le…
cs.LG2025
Blending Imitation and Reinforcement Learning for Robust Policy Improvement
Xuefeng Liu, Takuma Yoneda, Rick L. Stevens +2
While reinforcement learning (RL) has shown promising performance, its sample complexity continues to be a substantial hurdle, restricting its broader application across a variety…
cs.LG2025
DrugImproverGPT: A Large Language Model for Drug Optimization with Fine-Tuning via Structured Policy Optimization
Xuefeng Liu, Songhao Jiang, Siyu Chen +4
Finetuning a Large Language Model (LLM) is crucial for generating results towards specific objectives. This research delves into the realm of drug optimization and introduce a nove…