3 citations · 3 across the 8 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Adaptive Loss Balancing for Noise-Robust GRPO in Generative Recommendation
Kewei Xu, Junbo Qi, Yanyan Zou +3
Reinforcement learning (RL) presents a promising avenue for enhancing generative recommendation beyond supervised imitation, leveraging reward signals to guide policy improvement.…
cs.LG2025
Pelican-VL 1.0: A Foundation Brain Model for Embodied Intelligence
Yi Zhang, Che Liu, Xiancong Ren +20
This report presents Pelican-VL 1.0, a new family of open-source embodied brain models with parameter scales ranging from 7 billion to 72 billion. Our explicit mission is clearly s…