Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
ModelLens: Finding the Best for Your Task from Myriads of Models
Rui Cai, Weijie Jacky Mo, Xiaofei Wen +5
The open-source model ecosystem now contains hundreds of thousands of pretrained models, yet picking the best model for a new dataset is increasingly infeasible: new models and unb…
cs.LG2025
Learning More with Less: A Dynamic Dual-Level Down-Sampling Framework for Efficient Policy Optimization
Chao Wang, Tao Yang, Hongtao Tian +5
Critic-free methods like GRPO reduce memory demands by estimating advantages from multiple rollouts but tend to converge slowly, as critical learning signals are diluted by an abun…