INDEX
Explanations
phrases related to effort and commitment
New Auto-Interp
Negative Logits
ledge
-0.19
055
-0.16
whe
-0.15
дог
-0.14
Stopwatch
-0.14
Ferguson
-0.14
WithOptions
-0.14
wheel
-0.14
ormal
-0.14
å¯¾å¿ľ
-0.14
POSITIVE LOGITS
effort
0.21
.Private
0.18
cker
0.17
rello
0.16
zell
0.16
finishing
0.16
åĬŁ
0.15
isposable
0.15
ocrat
0.15
Scalars
0.15
Activations Density 0.024%