6 papers
Cross-Domain Energy-Guided Diffusion Generation for Off-Dynamics Reinforcement Learning
Yu Yang, Yihong Guo, Anqi Liu +1
Off-dynamics offline reinforcement learning seeks to learn a target-domain policy from a large source dataset and a limited target dataset under mismatched transition dynamics. Exi…
VeraRetouch: A Lightweight Fully Differentiable Framework for Multi-Task Reasoning Photo Retouching
Yihong Guo, Youwei Lyu, Jiajun Tang +5
Reasoning photo retouching has gained significant traction, requiring models to analyze image defects, give reasoning processes, and execute precise retouching enhancements. Howeve…
MOBODY: Model Based Off-Dynamics Offline Reinforcement Learning
Yihong Guo, Yu Yang, Pan Xu +1
We study off-dynamics offline reinforcement learning, where the goal is to learn a policy from offline source and limited target datasets with mismatched dynamics. Existing methods…
CorrectionPlanner: Self-Correction Planner with Reinforcement Learning in Autonomous Driving
Yihong Guo, Dongqiangzi Ye, Sijia Chen +2
Autonomous driving requires safe planning, but most learning-based planners lack explicit self-correction ability: once an unsafe action is proposed, there is no mechanism to corre…
Group-Sensitive Offline Contextual Bandits
Yihong Guo, Junjie Luo, Guodong Gao +2
Offline contextual bandits allow one to learn policies from historical/offline data without requiring online interaction. However, offline policy optimization that maximizes overal…
PAME-AI: Patient Messaging Creation and Optimization using Agentic AI
Junjie Luo, Yihong Guo, Anqi Liu +2
Messaging patients is a critical part of healthcare communication, helping to improve things like medication adherence and healthy behaviors. However, traditional mobile message de…