3 papers
cs.LG2026
On the Plasticity and Stability for Post-Training Large Language Models
Wenwen Qiang, Ziyin Gu, Jiahuan Zhou +4
Training stability remains a critical bottleneck for Group Relative Policy Optimization (GRPO), often manifesting as a trade-off between reasoning plasticity and general capability…
cs.CV2024
Self-Supervised Representation Learning with Meta Comprehensive Regularization
Huijie Guo, Ying Ba, Jie Hu +3
Self-Supervised Learning (SSL) methods harness the concept of semantic invariance by utilizing data augmentation strategies to produce similar representations for different deforma…
cs.LG2023
Rethinking Dimensional Rationale in Graph Contrastive Learning from Causal Perspective
Qirui Ji, Jiangmeng Li, Jie Hu +3
Graph contrastive learning is a general learning paradigm excelling at capturing invariant information from diverse perturbations in graphs. Recent works focus on exploring the str…