◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Boyi Wei

Princeton University

11 papers hereh-index 9814 citations12 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author7

Across the 11 of 11 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.CL3
  • cs.CR3
  • cs.LG2
affiliations
  • Princeton University
Homepage

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedYour Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents

1 citations · 1 across the 3 of their papers we have counts for

collaborators
Showing cs.CLShow all

3 papers · 1 filter

cs.CL2026

Large Language Models Generate Harmful Responses Using a Distinct Mechanism, Shared Across Harm Types

Hadas Orgad, Boyi Wei, Kaden Zheng +4

Large language models remain vulnerable to jailbreaks that elicit harmful responses, yet the mechanism behind harmful response generation is poorly understood. Here, we investigate…

cs.CL2026

Scaling Latent Reasoning via Looped Language Models

Rui-Jie Zhu, Zixuan Wang, Kai Hua +30

Modern LLMs are trained to "think" primarily via explicit text generation, such as chain-of-thought (CoT), which defers reasoning to post-training and under-leverages pre-training…

cs.CL2024

Evaluating Copyright Takedown Methods for Language Models

Boyi Wei, Weijia Shi, Yangsibo Huang +5

Language models (LMs) derive their capabilities from extensive training on diverse data, including potentially copyrighted material. These models can memorize and generate content…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.