◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Qiufeng Wang

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

fields
  • cs.AI1
  • cs.CL1
  • cs.CR1
ORCID 0000-0001-9385-7465
same name
  • Qiufeng Wang — 7 papers
  • Qiufeng Wang — 6 papers
  • Qiufeng Wang — 4 papers
  • Qiufeng Wang — 4 papers
  • Qiufeng Wang — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedTrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them

1 citations · 1 across the 3 of their papers we have counts for

collaborators

3 papers

cs.AI2025★ 1 cited

TrustJudge: Inconsistencies of LLM-as-a-Judge and How to Alleviate Them

Yidong Wang, Yunze Song, Tingyuan Zhu +11

The adoption of Large Language Models (LLMs) as automated evaluators (LLM-as-a-judge) has revealed critical inconsistencies in current evaluation frameworks. We identify two fundam…

cs.CL2025

Temporal Self-Rewarding Language Models: Decoupling Chosen-Rejected via Past-Future

Yidong Wang, Xin Wang, Cunxiang Wang +9

Self-Rewarding Language Models propose an architecture in which the Large Language Models(LLMs) both generates responses and evaluates its own outputs via LLM-as-a-Judge prompting,…

cs.CR2025

A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

Kun Wang, Guibin Zhang, Zhenhong Zhou +100

The remarkable success of Large Language Models (LLMs) has illuminated a promising pathway toward achieving Artificial General Intelligence for both academic and industrial communi…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.