◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Fabio Brau

7 papers hereh-index 16 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author7

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.AI2
  • cs.CR1
  • cs.CV1
same name
  • Fabio Brau — 2 papers, h 6

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.AIShow all

2 papers · 1 filter

cs.AI2026

Latent-space Attacks for Refusal Evasion in Language Models

Giorgio Piras, Raffaele Mura, Fabio Brau +4

Safety-aligned language models are trained to refuse harmful requests, yet refusal behavior can be suppressed by steering their internal representations. Existing methods do so by…

cs.AI2025

SOM Directions are Better than One: Multi-Directional Refusal Suppression in Language Models

Giorgio Piras, Raffaele Mura, Fabio Brau +3

Refusal refers to the functional behavior enabling safety-aligned language models to reject harmful or unethical prompts. Following the growing scientific interest in mechanistic i…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.