2 papers
cs.AI2025
Measuring Sparse Autoencoder Feature Sensitivity
Claire Tian, Katherine Tian, Nathan Hu
Sparse Autoencoder (SAE) features have become essential tools for mechanistic interpretability research. SAE features are typically characterized by examining their activating exam…
stat.ML2025
Counterfactual inference in sequential experiments
Raaz Dwivedi, Katherine Tian, Sabina Tomkins +3
We consider after-study statistical inference for sequentially designed experiments wherein multiple units are assigned treatments for multiple time points using treatment policies…