1 paper
Siqi Zeng, Andre N. Assis, Rowan Wang
We study self-modeling: an LLM's ability to answer questions about its own behavior. We focus on verifiable behavioral questions, such as whether a prompt edit would change the mod…