1 paper
Chunlin Tian, Kahou Tam, Yebo Wu +4
Deploying large language models (LLMs) in real-time systems remains challenging due to their substantial computational demands and privacy concerns. We propose Floe, a hybrid feder…