◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Nathan M. Truong

2 papers hereh-index 17 citations2 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • first author1

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.AI1
  • cs.CL1

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.CL2026

Abliteration Mitigation via Refusal Aliases

Nathan Truong

Abliteration, the removal of refusal capabilities from large language models by projecting weight matrices orthogonal to an extracted refusal direction, has emerged as a prominent…

cs.AI2026

LLM Scheming Inversely Scales with Pretraining Language Coverage

Nathan Truong, Aryan Panda, Rayming Ye +2

With the growing capabilities of frontier models, AI alignment becomes increasingly critical in high-risk deployment settings. While recent work has empirically demonstrated in-con…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.