© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Gemma-3-27B-IT
    3. 16-GEMMASCOPE-2-TRANSCODER-262K
    4. 101342
    Prev
    Next
    INDEX
    Explanations

    **Explaining Neuron Behavior**I'm currently focused on articulating the neuron's actions by correlating `TOP_POSITIVE_LOGITS`, `TOP_ACTIVATING_TEXTS`, `MAX_ACTIVATING_TOKENS`, and `TOKENS_AFTER_MAX_ACTIVATING_TOKEN` into a succinct descriptive phrase.Let's analyze the provided lists:1. **MAX_ACTIVATING_TOKENS**: * `cas` * `@` * `Z` * `ca` * `sizeof` (appears 3 times) * `holiday` * `nature` * `ab`2. **TOKENS_AFTER_MAX_ACTIVATING_TOKEN**: * `ins` (after `cas`) * `lock` (after `@`) * `azen` (after `Z`) * `/` (after `ca`) * `(` (after `sizeof` - appears 3 times) * `in` (after `holiday`) * `while` (after `nature`) * `uh` (after `ab`)3. **TOP_POSITIVE_LOGITS**: * ammonia * magnesium * poison * insects * amino * CH * quần * antioxidant * malware * poisonous4. **TOP_ACTIVATING_TEXTS**: * "summer and I want to know which shoes to wear. I will be wearing a linen suit. Can I wear driving moccasins or are they too casual?<end_of_turn> <start_of_turn>model Okay, a Crete wedding in a linen suit sounds *lovely" * "Lock _lock; public ReaderRelease(AsyncReaderWriterLock @lock) { _lock = @lock; } public void Dispose() { _lock.ReleaseReader();" * ", but don't get bogged down in intellectualizing them. The practice itself will reveal them. * **Zazen (Sitting Meditation):** The core practice. It's not about *stopping* thoughts, but about" * "0,6_IM118_KO7,21.htm](https://www.glassdoor.ca/Salaries/toronto-construction-worker-salary-SRCH_IL.0,6_IM1" (Contains `ca` followed by `[` and a URL; also `cas` seems related to `glassdoor.ca`) * "0, 20, 30}; int W = 50; int n = sizeof(values) / sizeof(values[0]); cout << knapsack(W, weights, values, n" (Repeated twice) * "0, 20, 30}; int W = 50; int n = sizeof(values) / sizeof(values[0]); cout << knapsack(W, weights, values, n" (Repeated twice) * "= {5, 2, 9, 1, 5, 6}; int size = sizeof(arr) / sizeof(arr[0]);" (Contains `sizeof`) * "<start_of_turn>user Give me some advice on a 6-day holiday in Tirana, Albania<end_of_turn> <start_of_turn>model Okay, here's a possible 6-day itinerary for Tir" (Contains `holiday`) * "It's a difficult subject, and I've tried to approach it with a tone that acknowledges the harshness of nature while avoiding gratuitous detail. **Please read the disclaimer at the end, as this is a sensitive topic.**" (Contains `nature`) * "able** **Headline:** Austin Entrepreneur with a Wanderlust Heart ✈️ **Summary:** "Mabuhay! 👋 I'm [Your Name], and I run a digital marketing company here in Austin, Texas" (Contains `ab` in `Austin`, `Mabuhay`)**Pattern Identification:**1. **`sizeof`**: Appears multiple times as a `MAX_ACTIVATING_TOKEN` and is followed by `(`. This strongly suggests programming or computational contexts, specifically array/type size calculation. The `TOP_POSITIVE_LOGITS` lists include `ammonia`, `magnesium`, `amino`, `CH`, which are chemical/elemental terms, but also `malware`, `poison`, `insects`, `poisonous`. This mix is confusing.2. **`holiday`**: Activating text is about planning a holiday. `TOKENS_AFTER_MAX_ACTIVATING_TOKEN` is `in`. This suggests 'holiday in X'.3. **`nature`**: Activating text discusses harshness of nature, sensitive topics. `TOKENS_AFTER_MAX_ACTIVATING_TOKEN` is `while`.4. **`@`**: Followed by `lock`. Suggests programming (`@lock` as a token/variable name).5. **`Z`**: Followed by `azen`. This points to the word "Zazen", a form of meditation, as seen in the activating text.6. **`cas` / `ca`**: In `TOP_ACTIVATING_TEXTS`, `glassdoor.ca` appears. `cas` could be the end of `.ca` or similar. The `TOP_POSITIVE_LOGITS` include `malware`, `poison`, `insects`, `poisonous`, and chemical terms. This might imply a domain or context where such terms are discussed, perhaps security or biology/health, but the `ca` strongly points to URLs.**Putting it Together:**The strongest recurring pattern is `sizeof` followed by `(`. This is a very specific programming construct.The tokens `cas`, `@`, and `ca` are also suggestive of code or web contexts.The `TOP_POSITIVE_LOGITS` include terms like `malware`, `poison`, `insects`, `poisonous`, `ammonia`, `magnesium`, `amino`, which can relate to technical/scientific domains where `sizeof` is used (e.g., data structures for biological simulations, security analysis of data types, etc.).The `holiday` and `nature` examples, while present, are less strongly represented in `MAX_ACTIVATING_TOKENS` compared to the technical/code-like ones. 'Zazen' is also specific but only one instance.The neuron seems to be firing in contexts involving technical enumeration or calculation, potentially in a domain that mixes programming with topics like security, biology/chemistry, or even general data handling where specific data structures (`sizeof`) are relevant.However, the 'poison', 'insects', 'malware' in `TOP_POSITIVE_LOGITS` alongside `sizeof` is a strong hint towards looking for contexts where these might be processed or discussed programmatically.Let's consider the words related to `sizeof` and what follows:`sizeof(values) / sizeof(values[0])` -> Calculating array length or size.`sizeof(arr) / sizeof(arr[0])` -> Same.This points to general computation on data structures.Now, let's re-examine the `TOP_POSITIVE_LOGITS`:`ammonia`, `magnesium`, `amino`, `CH` (chemical)`poison`, `insects`, `mal

    np_acts-logits-general · gemini-2.5-flash-lite
    New Auto-Interp
    Top Features by Cosine Similarity
    Configuration
    google/gemma-scope-2-27b-it/transcoder_all/layer_16_width_262k_l0_small_affine
    Prompts (Dashboard)
    238,145 prompts, 512 tokens each
    Dataset (Dashboard)
    lmsys + oasst1
    No Configuration Found
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
    ر
    0.52
    Gestion
    0.50
    represented
    0.49
    Neighborhood
    0.48
    Mounted
    0.48
    r
    0.47
    lukan
    0.47
    榄
    0.47
    Road
    0.46
    देशी
    0.46
    POSITIVE LOGITS
     ammonia
    0.56
     magnesium
    0.54
     poison
    0.53
     insects
    0.52
     amino
    0.51
     CH
    0.50
     quần
    0.50
     antioxidant
    0.49
     malware
    0.49
     poisonous
    0.49
    Activations Density 0.000%

    No Known Activations