INDEX
    Explanations

    phrases referring to the concept of 'first place' or 'initial conditions.'

    New Auto-Interp
    Negative Logits
    agit
    -0.17
     Placement
    -0.16
    away
    -0.14
    lein
    -0.14
    \db
    -0.14
    Placement
    -0.14
    chant
    -0.14
    ži
    -0.14
     mare
    -0.14
    ui
    -0.14
    POSITIVE LOGITS
     instance
    0.26
     place
    0.23
    instance
    0.20
    place
    0.20
     Instance
    0.19
    _instance
    0.18
    (instance
    0.17
    -instance
    0.17
     lugar
    0.17
    istance
    0.16
    Act Density 0.007%

    No Known Activations