◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tao Jiang

5 papers hereh-index 51.5k citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author2

Across the 3 of 5 papers where every author was matched, so the position is known.

fields
  • cs.AI2
  • cs.LG2
  • math.OC1
same name
  • Tao Jiang — 8 papers, h 9
  • Tao Jiang — 6 papers, h 5
  • Tao Jiang — 6 papers, h 2
  • Tao Jiang — 5 papers, h 3
  • Tao Jiang — 5 papers, h 2
  • Tao Jiang — 5 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.AI2026

CRAFT-GUI: Curriculum-Reinforced Agent For GUI Tasks

Songqin Nong, Xiaoxuan Tang, Jingxuan Xu +4

As autonomous agents become adept at understanding and interacting with graphical user interface (GUI) environments, a new era of automated task execution is emerging. Recent studi…

cs.LG2026

Kimi K2: Open Agentic Intelligence

Kimi Team, Yifan Bai, Yiping Bao +195

We introduce Kimi K2, a Mixture-of-Experts (MoE) large language model with 32 billion activated parameters and 1 trillion total parameters. We propose the MuonClip optimizer, which…

math.OC2025

Stochastic Approximation with Block Coordinate Optimal Stepsizes

Tao Jiang, Lin Xiao

We consider stochastic approximation with block-coordinate stepsizes and propose adaptive stepsize rules that aim to minimize the expected distance from the next iterate to an (unk…

cs.AI2025

Kimi k1.5: Scaling Reinforcement Learning with LLMs

Kimi Team, Angang Du, Bofei Gao +93

Language model pretraining with next token prediction has proved effective for scaling compute but is limited to the amount of available training data. Scaling reinforcement learni…

cs.LG2025

MoBA: Mixture of Block Attention for Long-Context LLMs

Enzhe Lu, Zhejun Jiang, Jingyuan Liu +22

Scaling the effective context length is essential for advancing large language models (LLMs) toward artificial general intelligence (AGI). However, the quadratic increase in comput…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.