◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Andreas Madsen

Mila

2 papers hereh-index 7579 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • middle author1

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.CL2
affiliations
  • Mila
HomepageORCID 0000-0002-1487-2796

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.CL2026

Scaling Inherently Interpretable Language Models

Guide Labs Team, Andreas Madsen, Aya Abdelsalam Ismail +7

Interpretability is often treated as a tax on capability: language models are trained as opaque systems, then explained after the fact, with methods whose reliability is difficult…

cs.CL2024

New Faithfulness-Centric Interpretability Paradigms for Natural Language Processing

Andreas Madsen

As machine learning becomes more widespread and is used in more critical applications, it's important to provide explanations for these models, to prevent unintended behavior. Unfo…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.