◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yifeng Gao

17 papers hereh-index 6157 citations20 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author15

Across the 17 of 17 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.CV4
  • cs.AI3
  • cs.CR3
  • cs.LG3
same name
  • Yifeng Gao — 5 papers, h 11
  • Yifeng Gao — 5 papers, h 2
  • Yifeng Gao — 5 papers, h 5
  • Yifeng Gao — 4 papers, h 1
  • Yifeng Gao — 3 papers, h 3
  • Yifeng Gao — 2 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedSafety at Scale: A Comprehensive Survey of Large Model and Agent Safety

1 citations · 2 across the 14 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

Towards Context-Invariant Safety Alignment for Large Language Models

Yixu Wang, Yang Yao, Xin Wang +4

Preference-based post-training aligns LLMs with human intent, yet safety behavior often remains brittle. A model may refuse a harmful request in a standard prompt but comply when t…

cs.CL2026

Internal Safety Collapse in Frontier Large Language Models

Yutao Wu, Xiao Liu, Yifeng Gao +7

This work identifies a critical failure mode in frontier large language models (LLMs), which we term Internal Safety Collapse (ISC): under certain task conditions, models enter a s…

cs.CL2025

ADMIT: Few-shot Knowledge Poisoning Attacks on RAG-based Fact Checking

Yutao Wu, Xiao Liu, Yinghui Li +5

Knowledge poisoning poses a critical threat to Retrieval-Augmented Generation (RAG) systems by injecting adversarial content into knowledge bases, tricking Large Language Models (L…

cs.CL2025

Identity Lock: Locking API Fine-tuned LLMs With Identity-based Wake Words

Hongyu Su, Yifeng Gao, Yifan Ding +1

The rapid advancement of Large Language Models (LLMs) has increased the complexity and cost of fine-tuning, leading to the adoption of API-based fine-tuning as a simpler and more e…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.