2 citations · 2 across the 23 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
CoRe-MoE: Compact Reusable MoE for Continual Multimodal Instruction Tuning
Runze Liu, Naibin Gu, Mingxu Ai +4
Continual multimodal instruction tuning requires multimodal large language models to acquire new task abilities sequentially while preserving previously learned knowledge. LoRA-MoE…
cs.AI2026
KnowRL: Boosting LLM Reasoning via Reinforcement Learning with Minimal-Sufficient Knowledge Guidance
Linhao Yu, Tianmeng Yang, Siyu Ding +8
RLVR improves reasoning in large language models, but its effectiveness is often limited by severe reward sparsity on hard problems. Recent hint-based RL methods mitigate sparsity…