Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Collaborative Large Language Model Inference via Resource-Aware Parallel Speculative Decoding
Jungyeon Koh, Hyun Jong Yang
The growing demand for on-device large language model (LLM) inference highlights the need for efficient mobile edge computing (MEC) solutions, especially in resource-constrained se…
cs.LG2025
Faithful and Fast Influence Function via Advanced Sampling
Jungyeon Koh, Hyeonsu Lyu, Jonggyu Jang +1
How can we explain the influence of training data on black-box models? Influence functions (IFs) offer a post-hoc solution by utilizing gradients and Hessians. However, computing t…