◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Di Wang

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author1

Across the 1 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG3
ORCID 0000-0002-7992-7743
same name
  • Di Wang — 30 papers
  • Di Wang — 13 papers, h 15
  • Di Wang — 9 papers, h 10
  • Di Wang — 8 papers, h 7
  • Di Wang — 7 papers
  • Di Wang — 7 papers, h 10

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.LG2025

PersRM-R1: Enhance Personalized Reward Modeling with Reinforcement Learning

Mengdi Li, Guanqiao Chen, Xufeng Zhao +3

Reward models (RMs), which are central to existing post-training methods, aim to align LLM outputs with human values by providing feedback signals during fine-tuning. However, exis…

cs.LG2024

Provably Efficient Action-Manipulation Attack Against Continuous Reinforcement Learning

Zhi Luo, Xiyuan Yang, Pan Zhou +1

Manipulating the interaction trajectories between the intelligent agent and the environment can control the agent's training and behavior, exposing the potential vulnerabilities of…

cs.LG2024

Incremental Structure Discovery of Classification via Sequential Monte Carlo

Changze Huang, Di Wang

Gaussian Processes (GPs) provide a powerful framework for making predictions and understanding uncertainty for classification with kernels and Bayesian non-parametric learning. Bui…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.