© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Qwen3-1.7B
    3. 27-LLAMASCOPE-2-LORSA-16K-K64
    4. 15373
    Prev
    Next
    INDEX
    Explanations

    say stop

    unknown · unknown
    New Auto-Interp
    Top Features by Cosine Similarity
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
    _rho
    -20.88
    _war
    -18.38
     Inflate
    -18.25
    rique
    -17.50
    esion
    -17.25
    _armor
    -16.88
    TargetException
    -16.38
     TORT
    -16.38
    _variance
    -16.38
    olen
    -16.25
    POSITIVE LOGITS
    停
    34.25
    暂停
    32.50
    停止
    30.50
     stop
    28.38
     Stop
    28.38
    中断
    27.88
     halt
    27.75
     halted
    27.75
    停工
    27.25
    stop
    27.13
    Activations Density 0.387%

    No Known Activations