◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Wei Da

4 papers hereh-index 00 citations4 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.DC4

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.DC2026

RouteBalance: Fused Model Routing and Load Balancing for Heterogeneous LLM Serving

Wei Da, Evangelia Kalyvianaki

Heterogeneous LLM serving stacks split scheduling into two layers that optimize in isolation: model routers pick a model from quality and cost signals while ignoring instance load,…

cs.DC2026

LLM-Emu: Native Runtime Emulation of LLM Inference via Profile-Driven Sampling

Wei Da, Evangelia Kalyvianaki

Realistic evaluation of LLM serving systems requires online workloads, dynamic arrivals, queueing, and the serving engine's local scheduling for execution batching, but running suc…

cs.DC2025

Dodoor: Efficient Randomized Decentralized Scheduling with Load Caching for Heterogeneous Tasks and Clusters

Wei Da, Evangelia Kalyvianaki

This paper presents Dodoor, a randomized decentralized scheduler for heterogeneous clusters. Dodoor removes hot-path probing via batched cache refreshes and introduces a heterogene…

cs.DC2025

Astrolabe: Balancing Load in LLM Serving with Randomized Prediction-Guided Scheduling

Wei Da, Evangelia Kalyvianaki

This paper presents Astrolabe, a randomized prediction-guided scheduler for one-shot request dispatch in multi-instance large language model (LLM) serving. Astrolabe improves load…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.