2 papers
cs.LG2026
Semantic-Space Exploration and Exploitation in RLVR for LLM Reasoning
Fanding Huang, Guanbo Huang, Xiao Fan +7
Reinforcement Learning with Verifiable Rewards (RLVR) for LLM reasoning is often framed as balancing exploration and exploitation in action space, typically operationalized with to…
cs.LG2025
Generative molecule evolution using 3D pharmacophore for efficient Structure-Based Drug Design
Yi He, Ailun Wang, Zhi Wang +3
Recent advances in generative models, particularly diffusion and auto-regressive models, have revolutionized fields like computer vision and natural language processing. However, t…