◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tianyi Zhou

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author2

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.CL1
same name
  • Tianyi Zhou — 32 papers, h 11
  • Tianyi Zhou — 23 papers, h 33
  • Tianyi Zhou — 15 papers
  • Tianyi Zhou — 11 papers
  • Tianyi Zhou — 10 papers, h 7
  • Tianyi Zhou — 9 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedDeja Vu: Contextual Sparsity for Efficient LLMs at Inference Time

19 citations · 19 across the 2 of their papers we have counts for

collaborators

4 papers

cs.LG2024

Fourier Circuits in Neural Networks and Transformers: A Case Study of Modular Arithmetic with Multiple Inputs

Chenyang Li, Yingyu Liang, Zhenmei Shi +2

In the evolving landscape of machine learning, a pivotal challenge lies in deciphering the internal representations harnessed by neural networks and Transformers. Building on recen…

cs.LG2023

Fast Heavy Inner Product Identification Between Weights and Inputs in Neural Network Training

Lianke Qin, Saayan Mitra, Zhao Song +2

In this paper, we consider a heavy inner product identification problem, which generalizes the Light Bulb problem~(\cite{prr89}): Given two sets A⊂{−1,+1}d and $B \sub…

cs.LG2023★ 19 cited

Deja Vu: Contextual Sparsity for Efficient LLMs at Inference Time

Zichang Liu, Jue Wang, Tri Dao +8

Large language models (LLMs) with hundreds of billions of parameters have sparked a new wave of exciting AI applications. However, they are computationally expensive at inference t…

cs.CL2023

Why Softmax Attention Outperforms Linear Attention

Yichuan Deng, Zhao Song, Kaijun Yuan +1

Large transformer models have achieved state-of-the-art results in numerous natural language processing tasks. Among the pivotal components of the transformer architecture, the att…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.