◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Suraj Srinivas

Harvard University

12 papers hereh-index 182.6k citations46 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author11

Across the 12 of 12 papers where every author was matched, so the position is known.

fields
  • cs.LG11
  • cs.CL1
affiliations
  • Harvard University
Homepage

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CLShow all

1 paper · 1 filter

cs.CL2025

Certifying LLM Safety against Adversarial Prompting

Aounon Kumar, Chirag Agarwal, Suraj Srinivas +3

Large language models (LLMs) are vulnerable to adversarial attacks that add malicious tokens to an input prompt to bypass the safety guardrails of an LLM and cause it to produce ha…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.