◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Keyu Duan

5 papers hereh-index 335 citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author2

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.AI1
  • cs.CL1
  • cs.RO1
same name
  • Keyu Duan — 3 papers, h 7
  • Keyu Duan — 2 papers
  • Keyu Duan — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators

5 papers

cs.AI2026

In-Context Reinforcement Learning for Tool Use in Large Language Models

Yaoqi Ye, Yiran Zhao, Keyu Duan +4

While large language models (LLMs) exhibit strong reasoning abilities, their performance on complex tasks is often constrained by the limitations of their internal knowledge. A com…

cs.LG2025

GEM: A Gym for Agentic LLMs

Zichen Liu, Anya Sims, Keyu Duan +16

The training paradigm for large language models (LLMs) is moving from static datasets to experience-based learning, where agents acquire skills via interacting with complex environ…

cs.LG2025

Efficient Process Reward Model Training via Active Learning

Keyu Duan, Zichen Liu, Xin Mao +5

Process Reward Models (PRMs) provide step-level supervision to large language models (LLMs), but scaling up training data annotation remains challenging for both humans and LLMs. T…

cs.CL2025

Unnatural Languages Are Not Bugs but Features for LLMs

Keyu Duan, Yiran Zhao, Zhili Feng +9

Large Language Models (LLMs) have been observed to process non-human-readable text sequences, such as jailbreak prompts, often viewed as a bug for aligned LLMs. In this work, we pr…

cs.RO2024

Novelty-based Sample Reuse for Continuous Robotics Control

Ke Duan, Kai Yang, Houde Liu +1

In reinforcement learning, agents collect state information and rewards through environmental interactions, essential for policy refinement. This process is notably time-consuming,…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.