◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Qifan Wang

14 papers hereh-index 9347 citations20 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author14

Across the 14 of 14 papers where every author was matched, so the position is known.

fields
  • cs.CL6
  • cs.CV4
  • cs.AI3
  • cs.MA1
same name
  • Qifan Wang — 16 papers, h 9
  • Qifan Wang — 14 papers, h 18
  • Qifan Wang — 10 papers, h 5
  • Qifan Wang — 8 papers, h 6
  • Qifan Wang — 8 papers, h 11
  • Qifan Wang — 6 papers, h 8

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.AIShow all

3 papers · 1 filter

cs.AI2026

IR3: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking

Mohammad Beigi, Ming Jin, Junshan Zhang +3

Reinforcement Learning from Human Feedback (RLHF) enables powerful LLM alignment but can introduce reward hacking - models exploit spurious correlations in proxy rewards without ge…

cs.AI2026

Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking

Mohammad Beigi, Ming Jin, Junshan Zhang +2

Reinforcement Learning from Human Feedback (RLHF) remains vulnerable to reward hacking, where models exploit spurious correlations in learned reward models to achieve high scores w…

cs.AI2025

Sycophancy Mitigation Through Reinforcement Learning with Uncertainty-Aware Adaptive Reasoning Trajectories

Mohammad Beigi, Ying Shen, Parshin Shojaee +5

Despite the remarkable capabilities of large language models, current training paradigms inadvertently foster \textit{sycophancy}, i.e., the tendency of a model to agree with or re…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.