2 papers
cs.CL2026
Value-and-Structure Alignment for Routing-Consistent Quantization of Mixture-of-Experts Models
Hancheol Park, Geonho Lee, Tairen Piao +1
Mixture-of-Experts (MoE) models scale foundation models efficiently by activating only a subset of experts for each token, but their large number of expert parameters still makes q…
cs.CL2024
Assessing the Answerability of Queries in Retrieval-Augmented Code Generation
Geonmin Kim, Jaeyeon Kim, Hancheol Park +2
Thanks to unprecedented language understanding and generation capabilities of large language model (LLM), Retrieval-augmented Code Generation (RaCG) has recently been widely utiliz…