INDEX
    Explanations
    New Auto-Interp
    Negative Logits
    _go
    -0.07
     œ
    -0.07
    -0.06
    feof
    -0.06
    ्त
    -0.06
     Доб
    -0.06
    eson
    -0.06
    Alternate
    -0.06
    -0.06
     Romans
    -0.06
    POSITIVE LOGITS
     نیست
    0.08
     Methodist
    0.08
    mnt
    0.07
    horizontal
    0.06
     merger
    0.06
     mycket
    0.06
    ()+
    0.06
     withdrawing
    0.06
     waving
    0.06
     går
    0.06
    Act Density 0.007%

    No Known Activations