© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsinterp-engineNEWAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Gemma-3-27B-IT
    3. 41-GEMMASCOPE-2-RES-262K
    4. 31200
    Prev
    Next
    INDEX
    Explanations

    affected

    np_acts-logits-general · gemini-2.5-flash-lite
    New Auto-Interp
    Top Features by Cosine Similarity
    Configuration
    google/gemma-scope-2-27b-it/resid_post_all/layer_41_width_262k_l0_big
    Prompts (Dashboard)
    238,145 prompts, 512 tokens each
    Dataset (Dashboard)
    lmsys + oasst1
    No Configuration Found
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
     excitatory
    0.40
     প্রেরণ
    0.39
    ෝ
    0.38
     compliments
    0.38
    ുന്ന
    0.37
    hesive
    0.37
     évident
    0.37
    inous
    0.36
    ﻖ
    0.36
    мовір
    0.35
    POSITIVE LOGITS
     affected
    2.25
    affected
    2.05
     Affected
    1.97
     afectado
    1.84
    Affected
    1.83
     प्रभावित
    1.81
     afectados
    1.81
     afectadas
    1.80
     afectada
    1.74
     afect
    1.69
    Activations Density 0.003%

    No Known Activations