◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xiyan Fu

4 papers hereh-index 7127 citations15 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author1

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.LG2

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.LG2026

Representation-Aware Advantage Estimation: Your Reward Model Provides More Than A Scalar Output

Guozheng Li, Xiyan Fu, Yiwen Guo

Current reinforcement learning from human feedback (RLHF) methods primarily rely on scalar rewards from a trained reward model (RM). While effective, scalar rewards are often noisy…

cs.CL2026

When Languages Disagree: Self-Evolving Multilingual LLM Judges

Xiyan Fu, Wei Lu

Multilingual LLM-as-a-judge is widely used to evaluate model outputs across languages, but suffers from cross-lingual inconsistency (Fu and Liu, 2025). Existing methods typically t…

cs.LG2026

Reinforcement Learning for Compositional Generalization with Outcome-Level Optimization

Xiyan Fu, Wei Liu

Compositional generalization refers to correctly interpret novel combinations of known primitives, which remains a major challenge. Existing approaches often rely on supervised fin…

cs.CL2025

How Reliable is Multilingual LLM-as-a-Judge?

Xiyan Fu, Wei Liu

LLM-as-a-Judge has emerged as a popular evaluation strategy, where advanced large language models assess generation results in alignment with human instructions. While these models…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.