◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Dayang Liang

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.AI2
  • cs.LG2
ORCID 0000-0002-6181-6746

identity via Semantic Scholar / OpenAlex

activity
20232026
most citedEpisodic Reinforcement Learning with Expanded State-reward Space

2 citations · 2 across the 4 of their papers we have counts for

collaborators

4 papers

cs.AI2026

SAPO: Single-Rollout Autoregressive Policy Optimization for Agentic Reinforcement Learning

Dayang Liang, Lang Feng, Bo An +1

Agentic reinforcement learning (RL) has become a critical stage in the post-training of large language models. Existing critic-free, group-relative methods estimate policy advantag…

cs.AI2026

PlanPO: Group Planning-Aware Policy Optimization for Multi-Turn Agentic LLMs

Dayang Liang, Liyuan He, Xuan Feng +3

Group-relative policy optimization has emerged as a key paradigm for training agentic large language models (LLMs) on multi-turn interactive tasks. However, most existing variants…

cs.LG2024★ 2 cited

Episodic Reinforcement Learning with Expanded State-reward Space

Dayang Liang, Yaru Zhang, Yunlong Liu

Empowered by deep neural networks, deep reinforcement learning (DRL) has demonstrated tremendous empirical successes in various domains, including games, health care, and autonomou…

cs.LG2023

Sequential Action-Induced Invariant Representation for Reinforcement Learning

Dayang Liang, Qihang Chen, Yunlong Liu

How to accurately learn task-relevant state representations from high-dimensional observations with visual distractions is a realistic and challenging problem in visual reinforceme…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.