INDEX
Explanations
references to awards and recognitions in sports
New Auto-Interp
Negative Logits
under
-0.16
hã
-0.15
popul
-0.15
obs
-0.15
Reform
-0.15
ouro
-0.15
constr
-0.15
recre
-0.14
char
-0.14
pract
-0.14
POSITIVE LOGITS
ì
0.20
ò
0.19
anche
0.18
azioni
0.17
zione
0.17
etÃł
0.17
azione
0.17
icol
0.16
lung
0.16
lett
0.16
Activations Density 0.152%