◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Andres Saurez

3 papers hereh-index 13 citations3 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.CL1

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.LG2026

Circuit Fingerprints: How Answer Tokens Encode Their Geometrical Path

Andres Saurez, Neha Sengar, Dongsoo Har

Circuit discovery and activation steering in transformers have developed as separate research threads, yet both operate on the same representational space. Are they two views of th…

cs.LG2026

Why Linear Interpretability Works: Invariant Subspaces as a Result of Architectural Constraints

Andres Saurez, Yousung Lee, Dongsoo Har

Linear probes and sparse autoencoders consistently recover meaningful structure from transformer representations -- yet why should such simple methods succeed in deep, nonlinear sy…

cs.CL2025

Continuous Adversarial Text Representation Learning for Affective Recognition

Seungah Son, Andrez Saurez, Dongsoo Har

While pre-trained language models excel at semantic understanding, they often struggle to capture nuanced affective information critical for affective recognition tasks. To address…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.