INDEX
Explanations
occurrences of the name "Bu" and its variations in different contexts
New Auto-Interp
Negative Logits
pery
-0.17
forth
-0.16
egra
-0.16
raj
-0.16
_accessible
-0.15
оÑĢÑĭ
-0.15
landa
-0.14
avel
-0.14
igua
-0.14
eg
-0.14
POSITIVE LOGITS
ceph
0.19
apest
0.18
ovice
0.17
levard
0.17
æ´ŀ
0.17
ýt
0.16
bu
0.16
htub
0.16
ilde
0.16
á»ĵng
0.16
Activations Density 0.016%