◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yudong Li

9 papers hereh-index 7196 citations18 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author6
  • middle author2

Across the 8 of 9 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.CV3
  • cs.CY1
  • cs.SD1
same name
  • Yudong Li — 5 papers, h 1
  • Yudong Li — 4 papers, h 2
  • Yudong Li — 3 papers, h 3
  • Yudong Li — 3 papers, h 2
  • Yudong Li — 2 papers, h 3
  • Yudong Li — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20222026
most citedCSL: A Large-scale Chinese Scientific Literature Dataset

13 citations · 14 across the 8 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

KoCo: Conditioning Language Model Pre-training on Knowledge Coordinates

Yudong Li, Jiawei Cai, Linlin Shen

Standard Large Language Model (LLM) pre-training typically treats corpora as flattened token sequences, often overlooking the real-world context that humans naturally rely on to co…

cs.CL2024

Dynamic data sampler for cross-language transfer learning in large language models

Yudong Li, Yuhao Feng, Wen Zhou +4

Large Language Models (LLMs) have gained significant attention in the field of natural language processing (NLP) due to their wide range of applications. However, training LLMs for…

cs.CL2022

Multi-stage Distillation Framework for Cross-Lingual Semantic Similarity Matching

Kunbo Ding, Weijie Liu, Yuejian Fang +3

Previous studies have proved that cross-lingual knowledge distillation can significantly improve the performance of pre-trained models for cross-lingual similarity matching tasks.…

cs.CL2022★ 13 cited

CSL: A Large-scale Chinese Scientific Literature Dataset

Yudong Li, Yuqing Zhang, Zhe Zhao +4

Scientific literature serves as a high-quality corpus, supporting a lot of Natural Language Processing (NLP) research. However, existing datasets are centered around the English la…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.