INDEX
    Explanations
    New Auto-Interp
    Negative Logits
     buff
    -0.09
     Full
    -0.08
     kath
    -0.07
     petition
    -0.07
     leid
    -0.07
    .Full
    -0.07
     petitions
    -0.07
     qar
    -0.07
     oz
    -0.07
     rada
    -0.07
    POSITIVE LOGITS
     firmly
    0.11
     anchors
    0.09
    0.09
     anchored
    0.09
     rooted
    0.08
     tether
    0.08
    下来
    0.08
    现实
    0.08
     Anch
    0.08
    anchors
    0.08
    Act Density 0.013%

    No Known Activations