2 papers
cs.LG2026
R^3: Replay, Reflection, and Ranking Rewards for LLM Reinforcement Learning
Zhizheng Jiang, Kang Zhao, Weikai Xu +5
Large reasoning models (LRMs) aim to solve diverse and complex problems through structured reasoning. Recent advances in group-based policy optimization methods have shown promise…
q-bio.BM2025
Fast and Accurate Antibody Sequence Design via Structure Retrieval
Xingyi Zhang, Kun Xie, Ningqiao Huang +5
Recent advancements in protein design have leveraged diffusion models to generate structural scaffolds, followed by a process known as protein inverse folding, which involves seque…