11 citations · 14 across the 16 of their papers we have counts for
11 papers · 1 filter
AGPO: Adaptive Group Policy Optimization with Dual Statistical Feedback
Miaobo Hu, Shuhao Hu, Bokun Wang +5
Reinforcement learning improves LLM reasoning, but PPO/GRPO typically use fixed clipping and decoding temperature, which makes training brittle and tuning-heavy. We propose Adaptiv…
DARE: Diffusion Language Model Activation Reuse for Efficient Inference
Natalia Frumkin, Bokun Wang, Hung-Yueh Chiang +3
Diffusion Large Language Models (dLLMs) have emerged as a promising alternative to auto-regressive (AR) models, offering greater expressive capacity and potential for parallel gene…
A Geometry-Aware Efficient Algorithm for Compositional Entropic Risk Minimization
Xiyuan Wei, Linli Zhou, Bokun Wang +2
This paper studies optimization for a family of problems termed , in which each data's loss is formulated as a Log-Expectation-Ex…
Stochastic Momentum Methods for Non-smooth Non-Convex Finite-Sum Coupled Compositional Optimization
Xingyu Chen, Bokun Wang, Ming Yang +2
Finite-sum Coupled Compositional Optimization (FCCO), characterized by its coupled compositional objective structure, emerges as an important optimization paradigm for addressing a…
Stochastic Primal-Dual Double Block-Coordinate for Two-way Partial AUC Maximization
Linli Zhou, Bokun Wang, My T. Thai +1
Two-way partial AUC (TPAUC) is a critical performance metric for binary classification with imbalanced data, as it focuses on specific ranges of the true positive rate (TPR) and fa…
Discovering Global False Negatives On the Fly for Self-supervised Contrastive Learning
Vicente Balmaseda, Bokun Wang, Ching-Long Lin +1
In self-supervised contrastive learning, negative pairs are typically constructed using an anchor image and a sample drawn from the entire dataset, excluding the anchor. However, t…