◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Sandro Andric

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • eess.SY1

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.LG2025

Brain-Grounded Axes for Reading and Steering LLM States

Sandro Andric

Interpretability methods for large language models (LLMs) typically derive directions from textual supervision, which can lack external grounding. We propose using human brain acti…

cs.LG2025

Do Large Language Models Walk Their Talk? Measuring the Gap Between Implicit Associations, Self-Report, and Behavioral Altruism

Sandro Andric

We investigate whether Large Language Models (LLMs) exhibit altruistic tendencies, and critically, whether their implicit associations and self-reports predict actual altruistic be…

cs.LG2025

BlockCert: Certified Blockwise Extraction of Transformer Mechanisms

Sandro Andric

Mechanistic interpretability aspires to reverse-engineer neural networks into explicit algorithms, while model editing seeks to modify specific behaviours without retraining. Both…

eess.SY2025

An Exact Quantile-Energy Equality for Terminal Halfspaces in Linear-Gaussian Control with a Discrete-Time Companion, KL/Schrodinger Links, and High-Precision Validation

Sandro Andric

We prove an exact equality between the minimal quadratic control energy and the squared normal-quantile gap for terminal halfspaces in linear-Gaussian systems with additive control…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.