1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2025
R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO
Huanjin Yao, Qixiang Yin, Jingyi Zhang +8
In this work, we aim to incentivize the reasoning ability of Multimodal Large Language Models (MLLMs) via reinforcement learning (RL) and develop an effective approach that mitigat…
cs.LG2024
DiffPoGAN: Diffusion Policies with Generative Adversarial Networks for Offline Reinforcement Learning
Xuemin Hu, Shen Li, Yingfen Xu +2
Offline reinforcement learning (RL) can learn optimal policies from pre-collected offline datasets without interacting with the environment, but the sampled actions of the agent ca…
cs.LG2021★ 1 cited
Spatial-Temporal-Fusion BNN: Variational Bayesian Feature Layer
Shiye Lei, Zhuozhuo Tu, Leszek Rutkowski +4
Bayesian neural networks (BNNs) have become a principal approach to alleviate overconfident predictions in deep learning, but they often suffer from scaling issues due to a large n…