2 papers
cs.CL2026
PRISM: A Unified Framework for Post-Training LLMs Without Verifiable Rewards
Mukesh Ghimire, Aosong Feng, Liwen You +3
Current techniques for post-training Large Language Models (LLMs) rely either on costly human supervision or on external verifiers to boost performance on tasks such as mathematica…
cs.LG2024
Learning Disentangled Equivariant Representation for Explicitly Controllable 3D Molecule Generation
Haoran Liu, Youzhi Luo, Tianxiao Li +2
We consider the conditional generation of 3D drug-like molecules with \textit{explicit control} over molecular properties such as drug-like properties (e.g., Quantitative Estimate…