◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yufei Ding

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.DC2
  • cs.DB1
  • cs.LG1
same name
  • Yufei Ding — 16 papers, h 22
  • Yufei Ding — 15 papers
  • Yufei Ding — 13 papers
  • Yufei Ding — 10 papers
  • Yufei Ding — 3 papers
  • Yufei Ding — 3 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedWLB-LLM: Workload-Balanced 4D Parallelism for Large Language Model Training

1 citations · 1 across the 2 of their papers we have counts for

collaborators

4 papers

cs.LG2025

Yggdrasil: Bridging Dynamic Speculation and Static Runtime for Latency-Optimal Tree-Based LLM Decoding

Yue Guan, Changming Yu, Shihan Fang +8

Speculative decoding improves LLM inference by generating and verifying multiple tokens in parallel, but existing systems suffer from suboptimal performance due to a mismatch betwe…

cs.DB2025

HedraRAG: Coordinating LLM Generation and Database Retrieval in Heterogeneous RAG Serving

Zhengding Hu, Vibha Murthy, Zaifeng Pan +4

This paper addresses emerging system-level challenges in heterogeneous retrieval-augmented generation (RAG) serving, where complex multi-stage workflows and diverse request pattern…

cs.DC2025

KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows

Zaifeng Pan, Ajjkumar Patel, Zhengding Hu +6

Large language model (LLM) based agentic workflows have become a popular paradigm for coordinating multiple specialized agents to solve complex tasks. To improve serving efficiency…

cs.DC2025★ 1 cited

WLB-LLM: Workload-Balanced 4D Parallelism for Large Language Model Training

Zheng Wang, Anna Cai, Xinfeng Xie +9

In this work, we present WLB-LLM, a workLoad-balanced 4D parallelism for large language model training. We first thoroughly analyze the workload imbalance issue in LLM training and…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.