MATH · IN · MODELS

Phi

Microsoft

Papers

On the Mutual Influence of Gender and Occupation in LLM Representations (2025), Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning (2026), LLaVA Steering: Visual Instruction Tuning with 500x Fewer Parameters through Modality Linear Representation-Steering (2024), Unravelling the Mechanisms of Manipulating Numbers in Language Models (2025), Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries (2026), Laguerre Geometry for Interpreting Large Language Models (2026), Steering Language Model Refusal with Sparse Autoencoder Features (2024), Geometric Factual Recall in Transformers (2026), Language Models Are Implicitly Continuous (2025), Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers (2025), The Geometries of Truth Are Orthogonal Across Tasks (2026), What are you sinking? A geometric approach on attention sink (2025), When Reward Hacking Rebounds: Understanding and Mitigating It with Representation-Level Signals (2026), Scale Determines Whether Language Models Organize Representation Geometry for Prediction (2026)