1 citations · 1 across the 5 of their papers we have counts for
Showing 2025 · cs.LGShow all
3 papers · 2 filters
cs.LG2025
Vicinity-Guided Discriminative Latent Diffusion for Privacy-Preserving Domain Adaptation
Jing Wang, Wonho Bae, Jiahong Chen +2
Recent work on latent diffusion models (LDMs) has focused almost exclusively on generative tasks, leaving their potential for discriminative transfer largely unexplored. We introdu…
cs.LG2025★ 1 cited
AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training
Zhenyu Han, Ansheng You, Haibo Wang +16
Reinforcement learning (RL) has become a pivotal technology in the post-training phase of large language models (LLMs). Traditional task-collocated RL frameworks suffer from signif…
cs.LG2025
A Survey of Optimization Methods for Training DL Models: Theoretical Perspective on Convergence and Generalization
Jing Wang, Anna Choromanska
As data sets grow in size and complexity, it is becoming more difficult to pull useful features from them using hand-crafted feature extractors. For this reason, deep learning (DL)…