◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yan Zhang

4 papers hereh-index 352 citations9 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.CV1
  • cs.LG1
same name
  • Yan Zhang — 54 papers, h 24
  • Yan Zhang — 39 papers, h 76
  • Yan Zhang — 23 papers
  • Yan Zhang — 15 papers, h 18
  • Yan Zhang — 15 papers
  • Yan Zhang — 14 papers, h 7

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20232025
most citedTraining and inference of large language models using 8-bit floating point

4 citations · 4 across the 3 of their papers we have counts for

collaborators

4 papers

cs.CV2025

MAGE: Multimodal Alignment and Generation Enhancement via Bridging Visual and Semantic Spaces

Shaojun E, Yuchen Yang, Jiaheng Wu +3

In the latest advancements in multimodal learning, effectively addressing the spatial and semantic losses of visual data after encoding remains a critical challenge. This is becaus…

cs.CL2024

DynaThink: Fast or Slow? A Dynamic Decision-Making Framework for Large Language Models

Jiabao Pan, Yan Zhang, Chen Zhang +3

Large language models (LLMs) have demonstrated emergent capabilities across diverse reasoning tasks via popular Chains-of-Thought (COT) prompting. However, such a simple and fast C…

cs.CL2024

Retrieval Augmented Instruction Tuning for Open NER with Large Language Models

Tingyu Xie, Jian Zhang, Yan Zhang +3

The strong capability of large language models (LLMs) has been applied to information extraction (IE) through either retrieval augmented prompting or instruction tuning (IT). Howev…

cs.LG2023★ 4 cited

Training and inference of large language models using 8-bit floating point

Sergio P. Perez, Yan Zhang, James Briggs +6

FP8 formats are gaining popularity to boost the computational efficiency for training and inference of large deep learning models. Their main challenge is that a careful choice of…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.