INDEX
    Explanations
    New Auto-Interp
    Negative Logits
     stems
    -0.07
     flu
    -0.06
     skiing
    -0.06
     Hill
    -0.06
    _LAYOUT
    -0.06
     Presbyterian
    -0.06
    _day
    -0.06
     April
    -0.06
     Book
    -0.06
     Separator
    -0.06
    POSITIVE LOGITS
     Na
    0.08
    μα
    0.07
    .autoconfigure
    0.07
    Na
    0.07
     naï
    0.07
    hydr
    0.07
     native
    0.06
    #+
    0.06
     Puerto
    0.06
    0.06
    Act Density 0.015%

    No Known Activations