41 citations · 115 across the 52 of their papers we have counts for
4 papers · 1 filter
From Training to Deployment: Post-Hoc Causal Feature Identification via Sensitivity Ratios
Athanasios Vlontzos, Giorgos Papanastasiou, Bernhard Kainz +1
Given a model that is already trained, which features does it rely on causally versus spuriously? Existing methods require access to the training procedure and cannot answer this p…
Entangled by Design: Spurious Intra-Variable Signal Routing in Tabular In-Context Learners
Athanasios Vlontzos, Giorgos Papanastasiou, Bernhard Kainz +1
Consider a model trained at a single hospital to predict patient recovery, where the measured feature bundles the patient's true health signal () with a systematic artefact…
Stress Testing Concept Erasure with Large Language Model Agents
Yuyang Xue, Feng Chen, Zhihua Liu +4
Concept erasure aims to remove semantic concepts from a trained generative model and is increasingly important for responsible AI deployment. However, verifying whether a model has…
CSEval: A Framework for Evaluating Clinical Semantics in Text-to-Image Generation
Robert Cronshaw, Konstantinos Vilouras, Junyu Yan +4
Text-to-image generation has been increasingly applied in medical domains for various purposes such as data augmentation and education. Evaluating the quality and clinical reliabil…