◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Brian Summa

4 papers hereh-index 211 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author2

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV2
  • cs.LG2

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.LG2026

Adaptive Multilevel Twisted Sequential Monte Carlo for Rare Events Estimation in Language Models

Zixuan Liu, Fangzheng Wu, Brian Summa +1

Rare unsafe behaviors in large language models can remain practically significant even when their probability is extremely small, particularly at deployment scales involving millio…

cs.LG2026

Robust General Utility for Reinforcement Learning

Zixuan Liu, Fangzheng Wu, Brian Summa +1

Reinforcement learning (RL) with general utility extends classic RL by optimizing an arbitrary utility functional of the policy-induced occupancy measure, thereby enabling a broade…

cs.CV2026

Attention Sinks in Diffusion Transformers: A Causal Analysis

Fangzheng Wu, Brian Summa

Attention sinks -- tokens that receive disproportionate attention mass -- are assumed to be functionally important in autoregressive language models, but their role in diffusion tr…

cs.CV2026

Model-Centric Diagnostics: A Framework for Internal State Readouts

Fangzheng Wu, Brian Summa

We present a model-centric diagnostic framework that treats training state as a latent variable and unifies a family of internal readouts -- head-gradient norms, confidence, entrop…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.