The neuron fires almost exclusively on those single, abnormally huge activation spikes—i.e. it isn’t picking out any normal word or feature at all but rather the rare overflow/overflow-style internal values. In other words, it flags those anomalously large internal activations rather than any meaningful token pattern.