1 paper
Su-Hyeon Kim, Jiwan Mun, Yo-Sub Han
Activation-based tools are usually tied to one model's native hidden space, requiring probes, sparse autoencoders, and natural-language interpreters to be rebuilt or rediscovered f…