Affine maps transfer SAEs, probes, and steering across model sizes
measured in 1 paperChen, Merullo, Stolfo & Pavlick fit affine maps between residual streams of differently-sized pretrained models and transfer whole SAEs, probes, and steering vectors [chen-etal-2025-transferring-linear-features-across-language-models-with-model-stitching] For the GPT-2 pair (small to medium) stitching preserves cross-entropy within roughly 9-11% overhead [chen-etal-2025-transferring-linear-features-across-language-models-with-model-stitching] Overhead is larger for other families (Pythia deduped 17-27%, Gemma-2 8.3-39%), so the 9-11% figure is GPT-2-specific [chen-etal-2025-transferring-linear-features-across-language-models-with-model-stitching] Using a transferred SAE as initialization for a larger target model cuts SAE training cost by roughly 50% [chen-etal-2025-transferring-linear-features-across-language-models-with-model-stitching]