◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Binhang Yuan

4 papers hereh-index 558 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.DC2
same name
  • Binhang Yuan — 20 papers, h 8
  • Binhang Yuan — 10 papers, h 7
  • Binhang Yuan — 5 papers, h 3
  • Binhang Yuan — 5 papers, h 3
  • Binhang Yuan — 3 papers, h 4
  • Binhang Yuan — 3 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators

4 papers

cs.DC2026

Schedule-Level Shared-Prefix Reuse for LLM RL Training

Pengbo Li, Feiyuan Zhang, Guangming Sheng +7

GRPO-based LLM post-training commonly samples multiple trajectories from the same prompt and then trains on the resulting group. In long-context GRPO workloads, this shared prompt-…

cs.CL2025

DeFT: Decoding with Flash Tree-attention for Efficient Tree-structured LLM Inference

Jinwei Yao, Kaiqi Chen, Kexun Zhang +4

Large language models (LLMs) are increasingly employed for complex tasks that process multiple generation calls in a tree structure with shared prefixes of tokens, including few-sh…

cs.CL2025

Locret: Enhancing Eviction in Long-Context LLM Inference with Trained Retaining Heads on Consumer-Grade Devices

Yuxiang Huang, Binhang Yuan, Xu Han +2

Scaling the input context length of a large language model (LLM) incurs a significant increase in computation cost and memory footprint to maintain the attention key-value (KV) cac…

cs.DC2024

LoHan: Low-Cost High-Performance Framework to Fine-Tune 100B Model on a Consumer GPU

Changyue Liao, Mo Sun, Zihan Yang +5

Nowadays, AI researchers become more and more interested in fine-tuning a pre-trained LLM, whose size has grown to up to over 100B parameters, for their downstream tasks. One appro…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.