© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsinterp-engineNEWAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Qwen3-1.7B
    3. 27-LLAMASCOPE-2-LORSA-16K-K64
    4. 13913
    Prev
    Next
    INDEX
    Explanations

    say benefit

    unknown · unknown
    New Auto-Interp
    Top Features by Cosine Similarity
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
     task
    -19.00
    输入
    -18.50
    任务
    -17.50
     input
    -16.50
    妻子
    -15.94
     tasks
    -15.50
     instructions
    -15.19
    的任务
    -15.06
    	task
    -15.00
     输入
    -14.94
    POSITIVE LOGITS
     benefits
    42.50
     benefit
    41.50
    benef
    39.50
     Benefits
    38.50
     Benefit
    36.75
    Benefits
    36.75
    益
    34.25
    受益
    34.25
    Benef
    34.00
     benefiting
    33.75
    Activations Density 0.694%

    No Known Activations