◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Daniel A. Herrmann

4 papers hereh-index 6231 citations15 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author1

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.LG1

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.LG2026

Benign interpolation and Occam's razor

Tom F. Sterkenburg, Daniel A. Herrmann, Jan-Willem Romeijn

Contemporary deep learning methods generalize well even when they fit their training data perfectly, a phenomenon known as benign interpolation. This phenomenon cannot be accounted…

cs.AI2026

Radical AI Interpretability

Daniel A. Herrmann, Benjamin A. Levinstein

We develop a framework for interpreting AI systems as agents, drawing on the philosophical tradition of radical interpretation and the tools of mechanistic interpretability. The co…

cs.AI2025

A Decision-Theoretic Approach for Managing Misalignment

Daniel A. Herrmann, Abinav Chari, Isabelle Qian +2

When should we delegate decisions to AI systems? While the value alignment literature has developed techniques for shaping AI values, less attention has been paid to how to determi…

cs.AI2025

Standards for Belief Representations in LLMs

Daniel A. Herrmann, Benjamin A. Levinstein

As large language models (LLMs) continue to demonstrate remarkable abilities across various domains, computer scientists are developing methods to understand their cognitive proces…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.