◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xiaoyan Cao

5 papers hereh-index 485 citations13 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author1

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.LG1
  • math.OC1

identity via Semantic Scholar / OpenAlex

activity
20212026
most citedRobust Model-based Reinforcement Learning for Autonomous Greenhouse Control

20 citations · 20 across the 4 of their papers we have counts for

collaborators
Showing cs.AIShow all

3 papers · 1 filter

cs.AI2026

Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems

Xiaoyang Cao, Siddarth Srinivasan, Michiel A. Bakker

End-to-end reinforcement learning can improve the accuracy of compound LLM systems, but it does not constrain how modules divide labor internally. We identify Role Drift, a failure…

cs.AI2025

RE-PO: Robust Enhanced Policy Optimization as a General Framework for LLM Alignment

Xiaoyang Cao, Zelai Xu, Mo Guang +4

Standard human preference-based alignment methods, such as Reinforcement Learning from Human Feedback (RLHF), are a cornerstone for aligning large language models (LLMs) with human…

cs.AI2021★ 20 cited

Robust Model-based Reinforcement Learning for Autonomous Greenhouse Control

Wanpeng Zhang, Xiaoyan Cao, Yao Yao +3

Due to the high efficiency and less weather dependency, autonomous greenhouses provide an ideal solution to meet the increasing demand for fresh food. However, managers are faced w…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.