INDEX
    Explanations
    New Auto-Interp
    Negative Logits
    Iterator
    -0.08
    ument
    -0.07
     가지고
    -0.07
     Battery
    -0.06
    Scheme
    -0.06
     Least
    -0.06
    Slash
    -0.06
     futuristic
    -0.06
    Verifier
    -0.06
    -0.06
    POSITIVE LOGITS
     kin
    0.14
     Kin
    0.12
    IN
    0.10
    in
    0.09
    Kin
    0.08
    .win
    0.07
    In
    0.07
     lin
    0.07
    kins
    0.07
    .MIN
    0.07
    Act Density 0.002%

    No Known Activations