Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Reference-guided Policy Optimization for Molecular Optimization via LLM Reasoning
Xuan Li, Zhanke Zhou, Zongze Li +4
Large language models (LLMs) benefit substantially from supervised fine-tuning (SFT) and reinforcement learning with verifiable rewards (RLVR) in reasoning tasks. However, these re…
cs.LG2025
The 1st International Workshop on Disentangled Representation Learning for Controllable Generation (DRL4Real): Methods and Results
Qiuyu Chen, Xin Jin, Yue Song +45
This paper reviews the 1st International Workshop on Disentangled Representation Learning for Controllable Generation (DRL4Real), held in conjunction with ICCV 2025. The workshop a…