SAE features from SDXL-Turbo's one-step U-Net transfer zero-shot and edit images
measured in 1 paperSurkov et al. train SAEs on transformer-block updates inside SDXL-Turbo's denoising U-Net [surkov-etal-2024-one-step-is-enough-sparse-autoencoders-for-text-to-image-diffusion-models] The resulting features generalize zero-shot, with no retraining, to 4-step SDXL-Turbo and the separately-trained multi-step SDXL-base model [surkov-etal-2024-one-step-is-enough-sparse-autoencoders-for-text-to-image-diffusion-models] Switching individual SAE features on or off during generation causally edits specific attributes of the generated image on the RIEBench benchmark [surkov-etal-2024-one-step-is-enough-sparse-autoencoders-for-text-to-image-diffusion-models] Different transformer blocks show measurable specialization by edit category [surkov-etal-2024-one-step-is-enough-sparse-autoencoders-for-text-to-image-diffusion-models]