1 paper
Xinjian Luo, Hongyan Chang, Jianxin Wei +5
Distributed large language model (LLM) inference frameworks connect isolated consumer-grade devices for large-scale model inference, substantially reducing hardware constraints. Ho…