2 citations · 2 across the 1 of their papers we have counts for
4 papers
Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space
Hengli Li, Chenxi Li, Tong Wu +8
Reasoning ability, a core component of human intelligence, continues to pose a significant challenge for Large Language Models (LLMs) in the pursuit of AGI. Although model performa…
Unlocking the Potential of Text-to-Image Diffusion with PAC-Bayesian Theory
Eric Hanchen Jiang, Yasi Zhang, Zhi Zhang +4
Text-to-image (T2I) diffusion models have revolutionized generative modeling by producing high-fidelity, diverse, and visually realistic images from textual prompts. Despite these…
Understanding Galaxy Morphology Evolution Through Cosmic Time via Redshift Conditioned Diffusion Models
Andrew Lizarraga, Eric Hanchen Jiang, Jacob Nowack +4
Redshift measures the distance to galaxies and underlies our understanding of the origin of the Universe and galaxy evolution. Spectroscopic redshift is the gold-standard method fo…
Statistical Guarantees for Lifelong Reinforcement Learning using PAC-Bayes Theory
Zhi Zhang, Chris Chow, Yasi Zhang +7
Lifelong reinforcement learning (RL) has been developed as a paradigm for extending single-task RL to more realistic, dynamic settings. In lifelong RL, the "life" of an RL agent is…