A function vector decomposes additively into causal per-shot sub-vectors
measured in 1 paperWang et al. redefine the function vector per-prompt and decompose it into additive per-example sub-vectors extracted via attention masking [wang-etal-2026-causal-decomposition-function-vectors] Global OLS fits weights so v_FV = sum_i w_i v_i + epsilon, reconstructing closely (mean cosine >=0.925, R^2 >=0.875) [wang-etal-2026-causal-decomposition-function-vectors] Mismatched-dictionary and orthogonalized-sub-FV null controls collapse to much lower fit, confirming the decomposition is not vacuous [wang-etal-2026-causal-decomposition-function-vectors] Injecting the reconstructed vector into 0-shot prompts recovers most of the full vector's causal steering effect (accuracy ratio 0.818-1.116) [wang-etal-2026-causal-decomposition-function-vectors] Under contextualization, attention shifts its share toward unambiguous examples (32% to 61%), with the query-key pathway dominating the gain [wang-etal-2026-causal-decomposition-function-vectors]