© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsinterp-engineNEWAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Qwen3-1.7B
    3. 26-LLAMASCOPE-2-LORSA-16K-K64
    4. 2445
    Prev
    Next
    INDEX
    Explanations

    say "aut" tokens

    unknown · unknown
    New Auto-Interp
    Top Features by Cosine Similarity
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
    uş
    -18.63
    istol
    -18.13
    爸爸妈妈
    -16.13
    éric
    -15.81
    Eric
    -15.75
    米尔
    -15.38
    eros
    -15.19
    erca
    -14.88
    奎
    -14.88
     Meredith
    -14.81
    POSITIVE LOGITS
     aut
    35.50
     Aut
    32.50
    Aut
    30.13
    mut
    27.00
    Mut
    26.88
     Mut
    26.38
     mut
    26.25
    aut
    26.00
    (mut
    25.88
    ut
    24.75
    Activations Density 0.079%

    No Known Activations