◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xiaoya Lu

12 papers hereh-index 6164 citations17 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author7

Across the 10 of 12 papers where every author was matched, so the position is known.

fields
  • cs.AI4
  • cs.CL3
  • cs.CR1
  • cs.CV1
  • cs.LG1
  • cs.MA1

identity via Semantic Scholar / OpenAlex

most citedLLMs know their vulnerabilities: Uncover Safety Gaps through Natural Distribution Shifts

1 citations · 1 across the 5 of their papers we have counts for

collaborators
Showing cs.CLShow all

3 papers · 1 filter

cs.CL2026

Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control

Lihao Sun, Lewen Yan, Xiaoya Lu +3

We show that emotion vectors in LLMs are organized by a two-dimensional valence-arousal (VA) subspace exhibiting circular geometry. Through principal component decomposition and ri…

cs.CL2026★ 1 cited

LLMs know their vulnerabilities: Uncover Safety Gaps through Natural Distribution Shifts

Qibing Ren, Hao Li, Dongrui Liu +7

Safety concerns in large language models (LLMs) have gained significant attention due to their exposure to potentially harmful data during pre-training. In this paper, we identify…

cs.CL2026

LLMs Deceive Unintentionally: Emergent Misalignment in Dishonesty from Misaligned Samples to Biased Human-AI Interactions

Xuhao Hu, Peng Wang, Xiaoya Lu +3

Previous research has shown that LLMs finetuned on malicious or incorrect completions within narrow domains (e.g., insecure code or incorrect medical advice) can become broadly mis…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.