◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ying-Xuan Zhuang

3 papers hereh-index 11 citations3 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.AI2
  • cs.LG1

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.AI2026

Dual-Axis Policy Optimization for LLM Agents: Bayesian Feedback Attribution and Trajectory Mass Normalization

Yingxuan Zhuang, Binhe Yu, Jingxiao Yang +6

Reinforcement learning for LLM agents involves two distinct optimization di- mensions: how environment feedback is exploited within a trajectory, and how complete trajectories are…

cs.AI2026

Reconciling Process Supervision with Outcome-Based Credit in Agentic Policy Optimization

Jingxiao Yang, Wangjie Gan, Yingxuan Zhuang +3

Outcome-based reinforcement learning provides verified feedback for language-model agents, but assigns trajectory-level advantage uniformly to all decisions, yielding coarse credit…

cs.LG2026

Mitigating Manifold Departure: Uncertainty-Aware Subspace Rectification for Trustworthy MLLM Decoding

Yingxuan Zhuang, Jingxiao Yang, Miao Pan +7

MLLMs frequently hallucinate objects inconsistent with visual inputs. This issue is typically attributed to the over-reliance on language priors, which can override the visual cont…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.