2 papers
cs.LG2026
TIDE: Temporal Incremental Draft Engine for Self-Improving LLM Inference
Jiyoung Park, Hankyu Jang, Changseok Song +1
Speculative decoding can substantially accelerate LLM inference, but realizing its benefits in practice is challenging due to evolving workloads and system-level constraints. We pr…
cs.CL2025
Expanding Foundational Language Capabilities in Open-Source LLMs through a Korean Case Study
Junghwan Lim, Gangwon Jo, Sungmin Lee +16
We introduce Llama-3-Motif, a language model consisting of 102 billion parameters, specifically designed to enhance Korean capabilities while retaining strong performance in Englis…