INDEX
    Explanations

    instances of the word "used" and its variants in various contexts

    New Auto-Interp
    Negative Logits
    _using
    -0.18
    using
    -0.18
    atur
    -0.17
    Using
    -0.17
     Using
    -0.17
     usando
    -0.16
     using
    -0.15
     USING
    -0.15
    -using
    -0.15
     usual
    -0.15
    POSITIVE LOGITS
    fully
    0.28
    full
    0.24
     interchange
    0.24
     sparing
    0.23
    ful
    0.23
     for
    0.23
     instead
    0.22
    age
    0.22
    fulness
    0.20
    FUL
    0.20
    Act Density 0.060%

    No Known Activations