3 papers
cs.CL2025
Masked-and-Reordered Self-Supervision for Reinforcement Learning from Verifiable Rewards
Zhen Wang, Zhifeng Gao, Guolin Ke
Test-time scaling has been shown to substantially improve large language models' (LLMs) mathematical reasoning. However, for a large portion of mathematical corpora, especially the…
cs.LG2024
A high-accuracy multi-model mixing retrosynthetic method
Shang Xiang, Lin Yao, Zhen Wang +4
The field of computer-aided synthesis planning (CASP) has seen rapid advancements in recent years, achieving significant progress across various algorithmic benchmarks. However, ch…
q-bio.BM2024
S-MolSearch: 3D Semi-supervised Contrastive Learning for Bioactive Molecule Search
Gengmo Zhou, Zhen Wang, Feng Yu +3
Virtual Screening is an essential technique in the early phases of drug discovery, aimed at identifying promising drug candidates from vast molecular libraries. Recently, ligand-ba…