◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Jacob Andreas

2 papers hereh-index 4126 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author1

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.CL2
same name
  • Jacob Andreas — 31 papers, h 53
  • Jacob Andreas — 16 papers
  • Jacob Andreas — 5 papers, h 7
  • Jacob Andreas — 5 papers, h 4
  • Jacob Andreas — 4 papers, h 2
  • Jacob Andreas — 4 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.CL2025

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users

Almog Hilel, Riddhi Bhagwat, Idan Shenfeld +2

We describe a vulnerability in language models (LMs) trained with user feedback, whereby a single user can persistently alter LM knowledge and behavior given only the ability to pr…

cs.CL2025

Can Gradient Descent Simulate Prompting?

Eric Zhang, Leshem Choshen, Jacob Andreas

There are two primary ways of incorporating new information into a language model (LM): changing its prompt or changing its parameters, e.g. via fine-tuning. Parameter updates incu…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.