INDEX
Explanations
words related to downplaying or minimizing something
words related to belief or disbelief
New Auto-Interp
Negative Logits
chnology
-0.76
office
-0.67
Detail
-0.66
Enhancement
-0.64
Methods
-0.63
LR
-0.61
OTA
-0.61
effects
-0.60
Fields
-0.60
zzo
-0.59
POSITIVE LOGITS
ieving
1.08
ievers
1.07
ief
1.04
iever
1.02
ichick
1.02
ieve
0.99
le
0.95
oved
0.95
ieved
0.93
ittle
0.90
Activations Density 0.029%