◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Robert Graham

3 papers hereh-index 245 citations3 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author2

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.CV2
  • cs.LG1

identity via Semantic Scholar / OpenAlex

works on
evaluation metrics 1generalization across domains 1ideological bias 1language model finetuning 1model alignment 1

From the 1 of 3 linked papers with an AI index.

collaborators

3 papers

cs.LG2026

Innocuous-Seeming Data, Latent Ideology: Ideological Generalisation in Finetuned LLMs

Robert Graham, Edward Stevinson, Yariv Barsheshat

The paper shows that fine‑tuning large language models on small, factually defensible datasets can cause broad ideological shifts across unrelated topics, and introduces metrics to…

cs.CV2025

Prisma: An Open Source Toolkit for Mechanistic Interpretability in Vision and Video

Sonia Joseph, Praneet Suresh, Lorenz Hufe +7

Robust tooling and publicly available pre-trained models have helped drive recent advances in mechanistic interpretability for language models. However, similar progress in vision…

cs.CV2025

Steering CLIP's vision transformer with sparse autoencoders

Sonia Joseph, Praneet Suresh, Ethan Goldfarb +6

While vision models are highly capable, their internal mechanisms remain poorly understood -- a challenge which sparse autoencoders (SAEs) have helped address in language, but whic…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.