◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Florian Tramèr

36 papers hereh-index 161.5k citations53 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author13
  • last author20

Across the 33 of 36 papers where every author was matched, so the position is known.

fields
  • cs.CR15
  • cs.LG10
  • cs.CL4
  • cs.CV3
  • cs.AI2
  • cs.CY1
same name
  • Florian Tramèr — 65 papers, h 54
  • Florian Tramèr — 4 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20222026
most citedJailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models

12 citations · 40 across the 34 of their papers we have counts for

collaborators
Showing 2026 · cs.CRShow all

4 papers · 2 filters

cs.CR2026

What Does It Mean to Break a Distillation Defense?

Lena Libon, Pura Peetathawatchai, Michael Aerni +2

Black-box LLMs (accessible only via API) are vulnerable to distillation attacks, in which an attacker queries the model and trains a student on its outputs. A recent line of work p…

cs.CR2026

Untrusted Content Masking for Web Agents with Security Guarantees

Kristina Nikolić, Egor Zverev, Javier Rando +3

Defenses that provide security guarantees against prompt injection attacks rely on strict isolation between trusted instructions and untrusted data. In text-based environments such…

cs.CR2026

Assessing Automated Prompt Injection Attacks in Agentic Environments

David Hofer, Edoardo Debenedetti, Florian Tramèr

Indirect prompt injection poses a critical threat to LLM agents that interact with untrusted external data, yet automated attack methods--proven effective for jailbreaking--remain…

cs.CR2026

Laundering AI Authority with Adversarial Examples

Jie Zhang, Pura Peetathawatchai, Florian Tramèr +1

Vision-language models (VLMs) are increasingly deployed as trusted authorities -- fact-checking images on social media, comparing products, and moderating content. Users implicitly…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.