1 paper
Robert Dahlke, Henrik Klagges, Dan Zecha +3
We present the Mixture-of-Tunable-Experts (MoTE), a method that extends the Mixture-of-Experts architecture of Large Language Models (LLMs). Without additional training, MoTE enabl…