◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

S. Pour

1 paper hereh-index 3341 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1

Across the 1 of 1 paper where every author was matched, so the position is known.

fields
  • cs.CL1

identity via Semantic Scholar / OpenAlex

most citedScalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

9 citations · 9 across the 1 of their papers we have counts for

collaborators

2 papers

cs.AI2024★ 44 cited

The AI risk repository: A meta-review, database, and taxonomy of risks from artificial intelligence

Peter Slattery, Alexander K. Saeri, Emily A. C. Grundy +7

Artificial intelligence (AI) is reshaping society, from video generation to medical diagnosis, coding agents to autonomous vehicles. Yet researchers, policymakers, and technology c…

cs.CL2023★ 9 cited

Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Rusheb Shah, Quentin Feuillade--Montixi, Soroush Pour +3

Despite efforts to align large language models to produce harmless responses, they are still vulnerable to jailbreak prompts that elicit unrestricted behaviour. In this work, we in…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.