21 citations · 48 across the 14 of their papers we have counts for
3 papers · 1 filter
Stress Testing Concept Erasure with Large Language Model Agents
Yuyang Xue, Feng Chen, Zhihua Liu +4
Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. However, verifying whether a model has…
A Causal Framework for Mitigating Data Shifts in Healthcare
Kurt Butler, Stephanie Riley, Damian Machlanski +13
Developing predictive models that perform reliably across diverse patient populations and heterogeneous environments is a core aim of medical research. However, generalization is o…
Optimal L2 Regularization in High-dimensional Continual Linear Regression
Gilad Karpel, Edward Moroshko, Ran Levinstein +3
We study generalization in an overparameterized continual linear regression setting, where a model is trained with L2 (isotropic) regularization across a sequence of tasks. We deri…