◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Run-Sheng Wang

29 papers hereh-index 10263 citations42 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author25
  • last author4

Across the 29 of 29 papers where every author was matched, so the position is known.

fields
  • cs.AR9
  • cs.CR7
  • cs.LG6
  • cs.CL3
  • cs.PF2
  • eess.AS1

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedAdapMoE: Adaptive Sensitivity-based Expert Gating and Management for Efficient MoE Inference

18 citations · 19 across the 24 of their papers we have counts for

collaborators
Showing cs.PFShow all

2 papers · 1 filter

cs.PF2025

HD-MoE: Hybrid and Dynamic Parallelism for Mixture-of-Expert LLMs with 3D Near-Memory Processing

Haochen Huang, Shuzhang Zhong, Zhe Zhang +5

Large Language Models (LLMs) with Mixture-of-Expert (MoE) architectures achieve superior model performance with reduced computation costs, but at the cost of high memory capacity a…

cs.PF2025

H2EAL: Hybrid-Bonding Architecture with Hybrid Sparse Attention for Efficient Long-Context LLM Inference

Zizhuo Fu, Xiaotian Guo, Wenxuan Zeng +6

Large language models (LLMs) have demonstrated remarkable proficiency in a wide range of natural language processing applications. However, the high energy and latency overhead ind…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.