MATH · IN · MODELS

Alpaca

Stanford

Structures found in this family (2)

By model (2)

Papers

Interpretability at Scale: Identifying Causal Mechanisms in Alpaca (2023), The Effectiveness of Style Vectors for Steering LLMs: A Human Evaluation (2026), Inference-Time Intervention: Eliciting Truthful Answers from a Language Model (2023)