1 citations · 1 across the 1 of their papers we have counts for
1 paper
Jaehan Kim, Minkyoo Song, Seungwon Shin +1
Recent large language models (LLMs) have increasingly adopted the Mixture-of-Experts (MoE) architecture for efficiency. MoE-based LLMs heavily depend on a superficial safety mechan…