3 citations · 5 across the 3 of their papers we have counts for
5 papers
Incast-Free MoE Rate-Based Scheduling
Evyatar Cohen, Jose Yallouz, Alexander Shpiner +3
Mixture of Experts (MoE) architectures have become key to large language models; however, their typical round-robin (RR) scheduling introduces significant bottlenecks. In this pape…
Scaling Routers with In-Package Optics and High-Bandwidth Memories
Isaac Keslassy, Ilay Yavlovich, Jose Yallouz +3
This paper aims to apply two major scaling transformations from the computing packaging industry to internet routers: the heterogeneous integration of high-bandwidth memories (HBMs…
Routing for Large ML Models
Ofir Cohen, Jose Yallouz Michael Schapira, Shahar Belkar +1
Training large language models (LLMs), and other large machine learning models, involves repeated communication of large volumes of data across a data center network. The communica…
Internet Performance in the 2022 Conflict in Ukraine: An Asymmetric Analysis
Tal Mizrahi, Jose Yallouz
On 24 February 2022 Russia invaded Ukraine, starting one of the largest military conflicts in Europe in recent years. In this paper we present preliminary findings about the impact…
Using Internet Measurements to Map the 2022 Ukrainian Refugee Crisis
Tal Mizrahi, Jose Yallouz
The conflict in Ukraine, starting in February 2022, began the largest refugee crisis in decades, with millions of Ukrainian refugees crossing the border to neighboring countries an…