3 papers
cs.LG2026
One step further with Monte-Carlo sampler to guide diffusion better
Minsi Ren, Wenhao Deng, Ruiqi Feng +1
Stochastic differential equation (SDE)-based generative models have achieved substantial progress in conditional generation via training-free differentiable loss-guided approaches.…
cs.LG2026
GenCP: Towards Generative Modeling Paradigm of Coupled Physics
Tianrun Gao, Haoren Zheng, Wenhao Deng +5
Real-world physical systems are inherently complex, often involving the coupling of multiple physics, making their simulation both highly valuable and challenging. Many mainstream…
cs.LG2025
Unlocking Reasoning Capabilities in LLMs via Reinforcement Learning Exploration
Wenhao Deng, Long Wei, Chenglei Yu +1
Reinforcement learning with verifiable rewards (RLVR) has recently enhanced the reasoning capabilities of large language models (LLMs), particularly for mathematical problem solvin…