© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Qwen3-1.7B
    3. 27-LLAMASCOPE-2-LORSA-16K-K64
    4. 15348
    Prev
    Next
    INDEX
    Explanations

    say "assistant"

    unknown · unknown
    New Auto-Interp
    Top Features by Cosine Similarity
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
    oce
    -18.00
    箭
    -17.00
    	del
    -16.50
    珠三角
    -16.50
    (del
    -16.38
     precip
    -16.13
    _Location
    -16.00
    德尔
    -16.00
    公开发行
    -16.00
    .locations
    -15.81
    POSITIVE LOGITS
     assistants
    18.88
    侍
    18.75
    助理
    17.50
    仆
    16.88
     assistant
    16.38
     aide
    16.00
     aides
    15.63
     loyal
    14.75
     Assistant
    14.44
    助手
    14.38
    Activations Density 0.105%

    No Known Activations