INDEX
    Explanations
    New Auto-Interp
    Negative Logits
    _dc
    -0.07
    _configuration
    -0.07
    DisplayName
    -0.06
    continuous
    -0.06
    knowledge
    -0.06
    -0.06
    Combo
    -0.06
     зад
    -0.06
     Müz
    -0.06
    x
    -0.06
    POSITIVE LOGITS
     falls
    0.11
     fall
    0.10
     Falls
    0.09
     fell
    0.09
     falling
    0.09
     fallen
    0.09
     Falling
    0.09
    fall
    0.07
     Fall
    0.07
     пад
    0.07
    Act Density 0.022%

    No Known Activations