© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsinterp-engineNEWAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Gemma-2-2B
    3. 0-CLT-HP
    4. 94351
    Prev
    Next
    INDEX
    Explanations

    apostrophe

    np_max-act · gemini-2.0-flash
    New Auto-Interp
    Top Features by Cosine Similarity
    Configuration
    Prompts (Dashboard)
    16,384 prompts, 128 tokens each
    Dataset (Dashboard)
    monology/pile-uncopyrighted
    No Configuration Found
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
    ?
    -1.30
    ?
    
    -1.10
    ?''
    -1.05
    ?」
    -1.03
    %?
    -1.00
    ?}
    -0.98
    ?");
    -0.96
    ?’
    -0.94
    ?...
    -0.92
    ?")
    -0.92
    POSITIVE LOGITS
     The
    0.74
    ↵
    0.72
     This
    0.71
     That
    0.70
     dignité
    0.66
    ↵↵
    0.65
     A
    0.65
     I
    0.63
     For
    0.63
     First
    0.62
    Activations Density 0.237%

    No Known Activations