◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Javier Rando

ETH Zurich

15 papers hereh-index 172.5k citations27 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author9
  • last author1

Across the 14 of 15 papers where every author was matched, so the position is known.

fields
  • cs.CR7
  • cs.LG3
  • cs.CL2
  • cs.CV2
  • cs.AI1
affiliations
  • ETH Zurich
HomepageORCID 0000-0002-2723-7660
same name
  • Javier Rando — 2 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedPosition: Adversarial ML for LLMs Is Not Making Any Progress

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing cs.AIShow all

1 paper · 1 filter

cs.AI2024

Universal Jailbreak Backdoors from Poisoned Human Feedback

Javier Rando, Florian Tramèr

Reinforcement Learning from Human Feedback (RLHF) is used to align large language models to produce helpful and harmless responses. Yet, prior work showed these models can be jailb…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.