Publications (99)
Identifying Policy Gradient Subspaces
Jan Schneider, Pierre Schumacher, Simon Guist +4
Policy gradient methods hold great potential for solving complex continuous control tasks. Still, their training efficiency can be improved by exploiting structure within the optim…
Comparison principle for stochastic heat equation on
Le Chen, Jingyu Huang
We establish the strong comparison principle and strict positivity of solutions to the following nonlinear stochastic heat equation on \[ \left(\frac{\partial }{\par…
OMPGPT: A Generative Pre-trained Transformer Model for OpenMP
Le Chen, Arijit Bhattacharjee, Nesreen Ahmed +4
Large language models (LLMs)such as ChatGPT have significantly advanced the field of Natural Language Processing (NLP). This trend led to the development of code-based large langua…
The third moment for the parabolic Anderson model
Le Chen
In this paper, we study the {\it parabolic Anderson model} starting from the Dirac delta initial data: \[ \left(\frac{\partial}{\partial t} -\fracν{2}\frac{\partial^2}{\partial x^…
Precise but Uncoupled: Reviewer Precision Does Not Guarantee Critique Uptake in Multi-Agent Math Reasoning
Chih-Hsuan Yang, Jingyan Jiang, Vikram Vasudevan +7
Many math- and science-oriented agent systems use hierarchical designs with specialized reviewer roles, assuming that a dedicated review stage should help turn wrong candidates int…
DataRaceBench V1.4.1 and DataRaceBench-ML V0.1: Benchmark Suites for Data Race Detection
Le Chen, Wenhao Wu, Stephen F. Siegel +2
Data races pose a significant threat in multi-threaded parallel applications due to their negative impact on program correctness. DataRaceBench, an open-source benchmark suite, is…