Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
3 results for “mechanistic interpretability”
Are Arithmetic Heuristic Neurons Form-Invariant? A Mechanistic Analysis of Symbols, Text, and Code in LLMs
A mechanistic interpretability study finds that arithmetic reasoning in Llama-3 models relies on a small, shared set of neurons across symbolic math, word problems, and Python code — suggesting failures stem from inconsistent activation states rather than format-specific circuitry.
Jul 21, 2026
Mechanistic interpretability researchers applying causality theory to LLMs
A Hacker News thread titled 'Mechanistic interpretability researchers applying causality theory to LLMs' contains user comments discussing early-stage academic efforts to use causal inference frameworks to understand internal mechanisms of large language models.
Jul 13, 2026
Representation as a Bottleneck for Mechanistic Interpretability: The Manifestation Unit Protocol
Researchers propose a new protocol to improve the reusability of neural-network component-level analyses.
Published Jul 2, 2026 · Analyzed Jul 5, 2026