MATH · IN · MODELS

Qwen

Alibaba

By model (75)

Qwen2.5-1.5B-Instruct · 1.5B
Qwen2-1.5B-Instruct · 1.5B
Qwen2-7B (base) · 7B
Qwen2.5-Coder-14B-Instruct · 14B
OpenR1-Qwen-7B · 7B
Qwen-1.8B-Chat · 1.8B
Qwen2.5-Math-1.5B · 1.5B
Qwen3-0.6B · 0.6B
Qwen3-1.7B · 1.7B
Qwen3-8B · 8B
Qwen 0.5B (base) · 0.5B
Qwen1.5-1.8B-Chat · 1.8B
Qwen1.5-32B-Chat · 32B
Qwen1.5-MoE-A2.7B · 14.3B total / 2.7B active, 60 experts/layer, top-4 routing
Qwen2-VL-7B-Instruct · 7B
Qwen2.5-0.5B · 0.5B
Qwen2.5-1.5B-Instruct · 1.5B
Qwen2.5-14B · 14B
Qwen2.5-VL 32B Instruct · 32B
Qwen2.5-VL-3B-Instruct · 3B
Qwen3-30B-A3B-Base · 30B-A3B
Qwen3-30B-A3B-Instruct · 30B
Qwen3-4B-Instruct-2507 · 4B
Qwen3-Embedding-0.6B · 0.6B
Qwen3-Embedding-8B · 8B
Qwen3-VL-235B-A22B-Instruct · 235B
Qwen3-VL 32B · 32B
Qwen3-VL-4B-Instruct · 4B
Qwen3-VL 8B · 8B
Qwen3-VL-8B-Instruct · 8B
Qwen3.5-35B-A3B (MoE, ~3B active/token) · 35B
Qwen3.5-4B · 4B
Qwen3.5-4B-Instruct · 4B
Qwen3.6-27B-Instruct · 27B
Qwen3.6-35B-A3B (MoE, 35B total / 3B active) · 35B-A3B
Qwen 7B (base) · 7B
QwQ-32B · 32B
Qwen-14B-Chat · 14B
no structures recorded for this checkpoint specifically
Qwen-72B-Chat · 72B
no structures recorded for this checkpoint specifically
Qwen-7B-Chat · 7B
no structures recorded for this checkpoint specifically
Qwen2.5-VL-72B-Instruct · 72B
no structures recorded for this checkpoint specifically
Qwen3-32B-Instruct · 32B
no structures recorded for this checkpoint specifically

Observations (131)

Papers

Geometry of Ordinal Representations in Language Models (2026), When Models Manipulate Manifolds: The Geometry of a Counting Task (2025), Semantic Structure of Feature Space in Large Language Models (2026), What Models Express, Suppress, and Resist: Auditing Open-Weight LLMs with Persona Vectors (2026), Death by a Thousand Directions: Exploring the Geometry of Harmfulness in LLMs through Subconcept Probing (2025), The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models (2026), The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models (2026), Understanding Moral Reasoning Trajectories in LLMs: Toward Probing-Based Explainability (2026), What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness (2026), Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering (2026), Tool-Call Dependency Structure Is Linearly Decodable in LLM Agent Residual Streams (2026), A Circuit for Predicting Hierarchical Structure In-Context in Large Language Models (2025), Cell-Based Representation of Relational Binding in Language Models (2026), Representational Analysis of Binding in Language Models (2024), Interpreting and Enhancing Emotional Circuits in Large Vision-Language Models via Cross-Modal Information Flow (2026), Tool Calling Is Linearly Readable and Steerable in Language Models (2026), Concept Heterogeneity-aware Representation Steering (2026), Quantitative Introspection in Language Models: Tracking Emotive States Across Conversation (2026), A Minimalist Structural Probe for Phase-Count Abstraction in Transformer Language Models (2026), Intrinsic Guardrails: How Semantic Geometry of Personality Interacts with Emergent Misalignment in LLMs (2026), A Geometric Account of Activation Steering through Angle-Norm Decomposition (2026), Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning (2026), Dimensional Collapse in Transformer Attention Outputs: A Challenge for Sparse Dictionary Learning (2025), Head Pursuit: Probing Attention Specialization in Multimodal Transformers (2025), Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization (2026), Conceptors for Semantic Steering (2026), The Geometry of Categorical and Hierarchical Concepts in Large Language Models (2024), When Language Representations Interact: Separability and Cross-Lingual Effects in LLMs (2026), A is for Absorption: Studying Feature Splitting and Absorption in Sparse Autoencoders (2024), Persona Vectors: Monitoring and Controlling Character Traits in Language Models (2025), Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinations (2026), Emergent Manifold Separability during Reasoning in Large Language Models (2026), Concepts Whisper While Syntax Shouts: Spectral Anti-Concentration and the Dual Geometry of Transformer Representations (2026), The Confidence Manifold: Geometric Structure of Correctness Representations in Language Models (2026), Language Models Represent and Transform Concepts with Shared Geometry (2026), Can LLMs Learn to Map the World from Local Descriptions? (2025), Unveiling the Latent Directions of Reflection in Large Language Models (2025), Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs (2025), The Dual Mechanisms of Spatial Reasoning in Vision-Language Models (2026), Scenario-based Probing and Steering Cultural Values in Large Language Models (2026), GEMS: Geometric Constraints Enable Multi-Semantic Superposition in LLMs (2026), Causal Probing for Internal Visual Representations in Multimodal Large Language Models (2026), A Shared Geometry of Difficulty in Multilingual Language Models (2026), Language Models Compare Quantities Using Number-specific and Unit-specific Heuristics (2026), The Geometry of Numerical Reasoning: Language Models Compare Numeric Properties in Linear Subspaces (2024), When Roleplaying, Do Models Believe What They Say? (2026), PERSONA: Dynamic and Compositional Inference-Time Personality Control via Activation Vector Algebra (2026), Shape Happens: Automatic Feature Manifold Discovery in LLMs via Supervised Multi-Dimensional Scaling (2025), Characterizing Linear Alignment Across Language Models (2026), How Does Alignment Tuning Shape Representations of Sycophancy and Related Cue-Induced Biases in LLMs? (2026), Latent Planning Emerges with Scale (2026), Reasoning Beyond Chain-of-Thought: A Latent Computational Mode in Large Language Models (2026), Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries (2026), Cross-model Transferability among Large Language Models on the Platonic Representations of Concepts (2025), Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations (2026), Global Evolutionary Steering: Refining Activation Steering Control via Cross-Layer Consistency (2026), Massive Values in Self-Attention Modules are the Key to Contextual Knowledge Understanding (2025), Beyond Steering Vector: Flow-based Activation Steering for Inference-Time Intervention (2026), Delta-Crosscoder: Robust Crosscoder Model Diffing in Narrow Fine-Tuning Regimes (2026), How Language Directions Align with Token Geometry in Multilingual LLMs (2026), Encoded but Not Routed: Explaining the Table-Chart Gap in Scientific Claim Verification (2026), Just-in-Time and Distributed Task Representations in Language Models (2025), Finding Lexical Identity and Inflectional Morphology in Modern Language Models (2025), What Really Controls Temporal Reasoning in LLMs: Tokenisation or Representation of Time? (2026), LLM Agents Already Know When to Call Tools — Even Without Reasoning (2026), Catching Rationalization in the Act: Detecting Motivated Reasoning Before and After CoT via Activation Probing (2026), Geometric Asymmetry in MoE Specialization: Functional Decorrelation and Representational Overlap (2026), Revealing Emergent Human-like Conceptual Representations from Language Prediction (2025), Number Representations in LLMs: A Computational Parallel to Human Perception (2025), Weber's Law in Transformer Magnitude Representations: Efficient Coding, Representational Geometry, and Psychophysical Laws in Language Models (2026), LLMs Know More About Numbers than They Can Say (2026), Relational Linearity is a Predictor of Hallucinations (2026), The Truthfulness Spectrum Hypothesis (2026), Shared Latent Structures Enable Unified Backdoor Detection and Mitigation in LLMs (2026), Invariant Reasoning Directions in Latent Trajectories of Language Models (2026), Attractor Geometry of Transformer Memory: From Conflict Arbitration to Confident Hallucination (2026), Why Far Looks Up: Probing Spatial Representation in Vision-Language Models (2026), Narrow Finetuning Leaves Clearly Readable Traces in Activation Differences (2026), Neural Chameleons: Language Models Can Learn to Hide Their Thoughts from Activation Monitors (2025), The Representational Geometry of Number (2026), Playing Devil's Advocate: Off-the-Shelf Persona Vectors Rival Targeted Steering for Sycophancy (2026), Trajectory Geometry of Transformer Representations Across Layers (2026), Understanding Subword Compositionality of Large Language Models (2025), Learning Uncertainty from Sequential Internal Dispersion in Large Language Models (2026), Steering at the Source: Style Modulation Heads for Robust Persona Control (2026), Probing then Editing Response Personality of Large Language Models (2025), Tracing Relational Knowledge Recall in Large Language Models (2026), Language Models Use Lookbacks to Track Beliefs (2025), Simulated Adoption: Decoupling Magnitude and Direction in LLM In-Context Conflict Resolution (2026), Geometric Factual Recall in Transformers (2026), The Geometry of Reasoning: Flowing Logics in Representation Space (2026), Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification (2025), Reasoning emerges from constrained inference manifolds in large language models (2026), The Geometry of Refusal in Large Language Models: Concept Cones and Representational Independence (2025), Refusal Direction is Universal Across Safety-Aligned Languages (2025), Refusal in Language Models Is Mediated by a Single Direction (2024), Representation Engineering: A Top-Down Approach to AI Transparency (2023), Programming Refusal with Conditional Activation Steering (2024), Fast Multi-dimensional Refusal Subspaces via RFM-AGOP (2026), Understanding and Preserving Safety in Fine-Tuned LLMs (2026), Linear Representations of Hierarchical Concepts in Language Models (2026), Latent Programming Horizons in Coding Agents (2026), Emergent Ordinal Geometry in Transformers Trained on Local Comparisons (2026), Convergent Linear Representations of Emergent Misalignment (2025), Linear Spatial World Models Emerge in Large Language Models (2025), Probing Spectrum-Like Organization of States of Mind in Transformer Representation Spaces (2026), On the Non-Identifiability of Steering Vectors in Large Language Models (2026), Steered LLM Activations Are Non-Surjective (2026), ReCoVeR the Target Language: Language Steering Without Sacrificing Task Performance (2025), Sycophancy Is Not One Thing: Causal Separation of Sycophantic Behaviors in LLMs (2025), Actionable Activation Directions for Detecting and Mitigating Emergent Misalignment Across Language Model Families (2026), Task Recognition and Task Learning Heads Align In-Context Hidden States with a Label-Unembedding Task Subspace (2026), The Geometries of Truth Are Orthogonal Across Tasks (2026), Task Vector Geometry Underlies Dual Modes of Task Inference in Transformers (2026), The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations (2026), Temporal Preference Concepts and Their Functions in a Large Language Model (2026), The Blessing and Curse of Dimensionality in Safety Alignment (2025), Anisotropy Decides Cosine vs. Rank Metrics for Text Embeddings (2026), Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight LMs (2026), Symmetry in Language Statistics Shapes the Geometry of Model Representations (2026), Not All Language Model Features Are One-Dimensionally Linear (2024), Do Sparse Autoencoders Capture Concept Manifolds? (2026), What are you sinking? A geometric approach on attention sink (2025), From Directions to Cones: Exploring Multidimensional Representations of Propositional Facts in LLMs (2025), How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs (2026), Valence-Arousal Subspace in LLMs: Circular Emotion Geometry and Multi-Behavioral Control (2026), Where Do Models Find Happiness? Emotion Vectors in Open-Source LLMs (2026), Negative Before Positive: Asymmetric Valence Processing in Large Language Models (2026), The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models (2026), Two Axes of LLM Abstention: Answer Correctness and Question Answerability (2026), Function-Vector Heads Are Two Populations: Writers and Cancellers in In-Context Learning (2026), The Myth of Expert Specialization in MoEs: Why Routing Reflects Geometry, Not Necessarily Domain Expertise (2026), Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models (2026), The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models (2026), Monitoring Emergent Reward Hacking During Generation via Internal Activations (2026), Rhetorical Questions in LLM Representations: A Linear Probing Study (2026), Behavioral Steering in a 35B MoE Language Model via SAE-Decoded Probe Vectors: One Agency Axis, Not Five Traits (2026), Spherical Steering: Geometry-Aware Activation Rotation for Language Models (2026), Decoding Emotion in the Deep: A Systematic Study of How LLMs Represent, Retain, and Express Emotion (2025)