◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Fan Lai

5 papers hereh-index 458 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author5

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.DC3
  • cs.DB1
  • cs.MA1
same name
  • Fan Lai — 7 papers, h 3
  • Fan Lai — 4 papers, h 3
  • Fan Lai — 4 papers, h 3
  • Fan Lai — 4 papers, h 13
  • Fan Lai — 2 papers, h 2
  • Fan Lai — 1 paper, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.DB2026

Compass: SLO-aware Query Planner for Compound AI Serving at Scale

Banruo Liu, Wei-Yu Lin, Minghao Fang +2

The rise of compound AI serving that integrates multiple operators in a pipeline enables end-user applications such as generative AI-powered meeting companions, autonomous driving,…

cs.DC2026

CacheFlow: Efficient LLM Serving with 3D-Parallel KV Cache Restoration

Sean Nian, Jiahao Fang, Qilong Feng +2

KV cache restoration has emerged as a dominant bottleneck in serving long-context LLM workloads, including multi-turn conversations, retrieval-augmented generation, and agentic pip…

cs.DC2025

JITServe: SLO-aware LLM Serving with Imprecise Request Information

Wei Zhang, Zhiyu Wu, Yi Mu +5

The integration of Large Language Models (LLMs) into applications ranging from interactive chatbots to multi-agent systems has introduced a wide spectrum of service-level objective…

cs.DC2025

Dora: QoE-Aware Hybrid Parallelism for Distributed Edge AI

Jianli Jin, Ziyang Lin, Qianli Dong +5

With the proliferation of edge AI applications, satisfying user quality of experience (QoE) requirements, such as model inference latency, has become a first class objective, as th…

cs.MA2025

Single-agent or Multi-agent Systems? Why Not Both?

Mingyan Gao, Yanzi Li, Banruo Liu +4

Multi-agent systems (MAS) decompose complex tasks and delegate subtasks to different large language model (LLM) agents and tools. Prior studies have reported the superior accuracy…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.