capability fusion 1domain adaptation 1inference-time routing 1large language models 1scientific question answering 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.CV2026
Relax Within, Balance Across: Geometry-Guided Load Balancing for Vision-Language Mixture-of-Experts
Ziang Wu, Peng Jin, Qishen Yin +4
Vision-language MoE batches contain different numbers of image and text tokens. Image resolution, image count, tiling, and prompt length all change this token mix. We call the stan…
cs.AI2026
Divergence Decoding: Training-Free Capability Fusion
Yimi Wang, Hao Li, Shuo Yang +6
The paper proposes Divergence Decoding, a training‑free method that dynamically routes token generation between a generalist LLM and a domain‑specialist LLM using Jensen‑Shannon di…