INDEX
Explanations
intensifiers, specifically the word "very" and similar expressions
New Auto-Interp
Negative Logits
oris
-0.15
ÑģÑĤÑİ
-0.15
embre
-0.15
оÑĢÑıд
-0.14
hti
-0.14
orre
-0.14
exter
-0.14
curacy
-0.14
ály
-0.14
\Migrations
-0.14
POSITIVE LOGITS
same
0.20
790
0.19
thing
0.18
same
0.17
SAME
0.17
essence
0.16
tons
0.16
ewan
0.16
588
0.16
opposite
0.15
Activations Density 0.018%