INDEX
Explanations
words related to the physical object "cup"
references to cups
New Auto-Interp
Negative Logits
targets
-0.68
Mes
-0.67
targeting
-0.66
targeted
-0.63
executed
-0.62
Tel
-0.61
Warn
-0.61
looting
-0.60
Shar
-0.60
serial
-0.59
POSITIVE LOGITS
cup
4.66
cup
1.34
cream
1.33
cups
1.25
bowl
1.24
cu
1.18
Cup
1.17
beer
1.15
Cups
1.06
pants
1.05
Activations Density 0.019%