◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Simon Yu

5 papers hereh-index 6122 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author4

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.AI1
  • cs.CR1
  • cs.LG1
same name
  • Simon Yu — 5 papers, h 0
  • Simon Yu — 2 papers, h 3
  • Simon Yu — 2 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.CR2026

Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Agents

Xu Li, Simon Yu, Minzhou Pan +5

LLM-based agents are becoming increasingly capable, yet their safety lags behind. This creates a gap between what agents can do and should do. This gap widens as agents engage in m…

cs.CL2026

PolySkill: Learning Generalizable Skills Through Polymorphic Abstraction

Simon Yu, Gang Li, Weiyan Shi +1

Large language models (LLMs) are moving beyond static uses and are now powering agents that learn continually during their interaction with external environments. For example, agen…

cs.AI2026

SPIRAL: Self-Play on Zero-Sum Games Incentivizes Reasoning via Multi-Agent Multi-Turn Reinforcement Learning

Bo Liu, Leon Guertler, Simon Yu +9

Recent advances in reinforcement learning have shown that language models can develop sophisticated reasoning through training on tasks with verifiable rewards, but these approache…

cs.LG2026

GEM: A Gym for Agentic LLMs

Zichen Liu, Anya Sims, Keyu Duan +16

The training paradigm for large language models (LLMs) is moving from static datasets to experience-based learning, where agents acquire skills via interacting with complex environ…

cs.CL2025

TextArena

Leon Guertler, Bobby Cheng, Simon Yu +3

TextArena is an open-source collection of competitive text-based games for training and evaluation of agentic behavior in Large Language Models (LLMs). It spans 57+ unique environm…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.