© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsinterp-engineNEWAPI Community BlogPrivacy & TermsContact
    EXPLANATION TYPE
    oai_token-act-pair
    Description
    OpenAI's Automated Interpretability from paper "Language models can explain neurons in language models". Modified by Johnny Lin to add new models/context windows.
    Author
    OpenAI
    URL
    https://github.com/hijohnnylin/automated-interpretability
    Settings
    Default prompts from the main branch, strategy TokenActivationPair. Uses top 10 deduplicated activations.
    Recent Explanations
    alphanumeric classification markers, especially bracketed numeral–letter locants and group/key labels indicating categorical or positional codes.
    gpt-5
    piece, 'Fart in C Major,' immediately sets
    Neuronpedia logo
    LLAMA3.3-70B-IT
    50-RESID-POST-GF
    INDEX 31374
    the subtoken “mi,” whether standing alone or embedded in names, abbreviations, and code identifiers.
    gpt-5
     (mi.getLockedStackDepth
    Neuronpedia logo
    GEMMA-2-2B
    12-GEMMASCOPE-RES-16K
    INDEX 8252
    decimal numbers with a decimal point (floating-point values) appearing in text.
    gpt-5
    2001.  The need for this criteria
    Neuronpedia logo
    GEMMA-2-2B
    12-GEMMASCOPE-RES-16K
    INDEX 2891
    the English function word used as a demonstrative or clause-introducing complementizer, especially when capitalized at the start of a sentence.
    gpt-5
     is a hard one. That
    Neuronpedia logo
    GEMMA-2-2B
    12-GEMMASCOPE-RES-16K
    INDEX 4608
    mentions of specific years, especially four-digit years in dates (notably from the 1900s).
    gpt-5
    8, 1999, Washington, D
    Neuronpedia logo
    GEMMA-2-2B
    12-GEMMASCOPE-RES-16K
    INDEX 14356
    enumerations of snack foods presented as example lists, especially following an introductory phrase and joined by commas and a conjunction.
    gpt-5
    treats, such as chips, candy, and cookies.\n
    Neuronpedia logo
    LLAMA3.3-70B-IT
    50-RESID-POST-GF
    INDEX 63166
    Sentences and phrases about personal growth, therapy, healing, and emotional support—self-help / counselling language focused on feelings, confidence, choices, and recovery.
    gpt-5-mini
    voice so that you can make clear decisions?↵↵Welcome!
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 2728
    The neuron detects vivid, poetic/lyrical language—emotionally charged descriptive imagery and rhetorical, metaphorical phrasing.
    gpt-5-mini
    home, you say But home isn't stony chunks
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 27900
    words and phrases expressing brokenness, loss, and painful/heartbroken memories.
    gpt-5-mini
    fromthose painful days.Lost in darkness from the painful
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 9317
    The neuron lights up on action and stance words—main verbs, auxiliaries, and negations that signal commands, judgments, or emphatic assertions.
    gpt-5-mini
    's done. But do not take the life of a
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 7228
    The neuron detects numeric year/date tokens (digits that mark years or dates in text).
    gpt-5-mini
    to compete at the 1968 Cannes Film
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 20775
    It detects the main topic or subject words in question titles/headers—concise nouns or technical keywords that name what the question is about.
    gpt-5-mini
    currency symbol is to the left or right of the price
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 37370
    words that signal strong subjective emphasis or novelty (emotive/adjective or adverbial hype like "new", "magically", "never-before-seen", or other high-intensity descriptors).
    gpt-5-mini
    collection. It will not magically add brand new, never
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 31218
    The neuron detects personal references — pronouns and possessives that point to people or groups (e.g., I, you, we, they, our, their, us).
    gpt-5-mini
    wellbeing, regardless of whether we may have a lived experience
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 21977
    sentences or fragments from legal, copyright, licensing, or policy/disclaimer text (formal legal wording).
    gpt-5-mini
    external organisation or party; nor to use behavioural analysis for
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 18382
    words describing coping behaviors, maladaptive responses, and symptoms related to trauma/mental health.
    gpt-5-mini
    other activities, and using humor. In contrast, female
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 49180
    Medical and health-related content — especially disease names, symptoms, treatments, and preventive-care information.
    gpt-5-mini
    is a test for the presence of a parasite or worm
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 62391
    The neuron detects explanatory/encyclopedic phrasing that describes what something is used for or how it functions (phrases about use, purpose, interaction, or tasks).
    gpt-5-mini
    interact with one another and with other systems to perform various
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 51972
    medical and biomedical terms—diseases, anatomical sites, proteins, pathogens, and related clinical/pathology vocabulary.
    gpt-5-mini
    by engineered cytochrome P450 enzymes.↵
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 57573
    Boilerplate legal/administrative text—section headings and sentences from terms, privacy, cookies, policies, and other formal notices.
    gpt-5-mini
    that person from further participation. So please, just remember
    Neuronpedia logo
    QWEN3-32B
    32-RESID-BATCHTOPK-65K
    INDEX 56177