Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
OThink-MR1: Stimulating multimodal generalized reasoning capabilities via dynamic reinforcement learning
Zhiyuan Liu, Yuting Zhang, Feng Liu +3
Multimodal Large Language Models (MLLMs) have gained significant traction for their ability to process diverse input data types and generate coherent, contextually relevant outputs…
cs.LG2024
Adaptive Conditional Expert Selection Network for Multi-domain Recommendation
Kuiyao Dong, Xingyu Lou, Feng Liu +4
Mixture-of-Experts (MOE) has recently become the de facto standard in Multi-domain recommendation (MDR) due to its powerful expressive ability. However, such MOE-based method typic…