2 citations · 3 across the 12 of their papers we have counts for
Showing 2026Show all
2 papers · 1 filter
cs.AI2026
Balancing Performance and Diversity in GRPO Autoregressive Text-to-Image Post-Training
Yuanhao Chiang, Hongbo Duan, Chunru Yang +3
Autoregressive text-to-image (T2I) generation has recently advanced rapidly, yet aligning generated images with human preferences remains challenging. GRPO-style online reinforceme…
cs.LG2026
UACER: An Uncertainty-Adaptive Critic Ensemble Framework for Robust Adversarial Reinforcement Learning
Jiaxi Wu, Tiantian Zhang, Yuxing Wang +2
Robust adversarial reinforcement learning has emerged as an effective paradigm for training agents to handle uncertain disturbance in real environments, with critical applications…