◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Sebastian Szyller

7 papers hereh-index 101.2k citations19 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3
  • last author2

Across the 5 of 7 papers where every author was matched, so the position is known.

fields
  • cs.CL3
  • cs.CR3
  • cs.CV1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CLShow all

3 papers · 1 filter

cs.CL2026

How Context Attribution Handles What the Model Already Knows

Quoc-Huy Trinh, Lin Zhu, Sebastian Szyller

Context attribution methods for large language models (LLMs) identify which input context contributes to the model response. Recent works show the initial success in attributing th…

cs.CL2025

Soft Token Attacks Cannot Reliably Audit Unlearning in Large Language Models

Haokun Chen, Sebastian Szyller, Weilin Xu +1

Large language models (LLMs) are trained using massive datasets, which often contain undesirable content such as harmful texts, personal information, and copyrighted material. To a…

cs.CL2024

LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked

Mansi Phute, Alec Helbling, Matthew Hull +4

Large language models (LLMs) are popular for high-quality text generation but can produce harmful content, even when aligned with human values through reinforcement learning. Adver…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.