1 citations · 2 across the 12 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
What Does it Mean for a Neural Network to Learn a "World Model"?
Kenneth Li, Fernanda Viégas, Martin Wattenberg
We propose a set of precise criteria for saying a neural net learns and uses a "world model." The goal is to give an operational meaning to terms that are often used informally, in…
cs.AI2025
The Geometry of Self-Verification in a Task-Specific Reasoning Model
Andrew Lee, Lihao Sun, Chris Wendler +2
How do reasoning models verify their own answers? We study this question by training a model using DeepSeek R1's recipe on the CountDown task. We leverage the fact that preference…
cs.AI2024★ 1 cited
Relational Composition in Neural Networks: A Survey and Call to Action
Martin Wattenberg, Fernanda B. Viégas
Many neural nets appear to represent data as linear combinations of "feature vectors." Algorithms for discovering these vectors have seen impressive recent success. However, we arg…