◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Pete Walsh

2 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.CL2

identity via Semantic Scholar / OpenAlex

most citedStaged Training for Transformer Language Models

4 citations · 5 across the 2 of their papers we have counts for

collaborators

2 papers

cs.CL2022★ 1 cited

Continued Pretraining for Better Zero- and Few-Shot Promptability

Zhaofeng Wu, Robert L. Logan, Pete Walsh +4

Recently introduced language model prompting methods can achieve high accuracy in zero- and few-shot settings while requiring few to no learned task-specific parameters. Neverthele…

cs.CL2022★ 4 cited

Staged Training for Transformer Language Models

Sheng Shen, Pete Walsh, Kurt Keutzer +3

The current standard approach to scaling transformer language models trains each model size from a different random initialization. As an alternative, we consider a staged training…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.