◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Qizhen Zhang

9 papers hereh-index 594 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author5

Across the 9 of 9 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • cs.CL3
  • cs.AI2
same name
  • Qizhen Zhang — 6 papers, h 3
  • Qizhen Zhang — 3 papers, h 15
  • Qizhen Zhang — 1 paper, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20212026
most citedBAM! Just Like That: Simple and Efficient Parameter Upcycling for Mixture of Experts

1 citations · 1 across the 8 of their papers we have counts for

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2026

Test-Time Adaptation via Many-Shot Prompting: Benefits, Limits, and Pitfalls

Shubhangi Upasani, Chen Wu, Jay Rainton +4

Test-time adaptation enables large language models (LLMs) to modify their behavior at inference without updating model parameters. A common approach is many-shot prompting, where l…

cs.LG2026

An Empirical Study on Noisy Data and LLM Pretraining Loss Divergence

Qizhen Zhang, Ankush Garg, Jakob Foerster +3

Large-scale pretraining datasets drive the success of large language models (LLMs). However, these web-scale corpora inevitably contain large amounts of noisy data due to unregulat…

cs.LG2024★ 1 cited

BAM! Just Like That: Simple and Efficient Parameter Upcycling for Mixture of Experts

Qizhen Zhang, Nikolas Gritsch, Dwaraknath Gnaneshwar +8

The Mixture of Experts (MoE) framework has become a popular architecture for large language models due to its superior performance over dense models. However, training MoEs from sc…

cs.LG2024

Analysing the Sample Complexity of Opponent Shaping

Kitty Fung, Qizhen Zhang, Chris Lu +3

Learning in general-sum games often yields collectively sub-optimal results. Addressing this, opponent shaping (OS) methods actively guide the learning processes of other agents, e…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.