© Neuronpedia 2026
    Privacy & TermsBlogGitHubSlackTwitterContact
    Neuronpedia logo - a computer chip with a rounded viewfinder border around it

    Neuronpedia

    Jacobian LensNEW
    Natural Language
    Autoencoders
    NEW
    Assistant AxisNEWCircuit TracerUPDATESteerSAE EvalsExportsAPI Community BlogPrivacy & TermsContact
    1. Home
    2. Olmo-3-1025-7B
    3. 16-RES-MATRYOSHKA-65K
    4. 1179
    Prev
    Next
    INDEX
    Explanations

    at java.The commonality seems to be `at` followed by a package or domain name, commonly seen in error messages or code call stacks. The positive logits suggest words like "unknown" or "anonymous", which could relate to the origin or lack of clear identification within these logs.Let's refine:MAX_ACTIVATING_TOKENS is dominated by 'at'.TOKENS_AFTER_MAX_ACTIVATING_TOKEN includes 'org', 'com'.TOP_POSITIVE_LOGITS includes 'unknown', 'anonymous'.TOP_ACTIVATING_TEXTS includes many lines starting with 'at' followed by package names/file paths, characteristic of call stacks.The neuron seems to detect `at` followed by organizational / package identifiers, often in error contexts. This aligns with `at` in call stacks.Phrase idea: "at" followed by domain/package names.Alternative: "at" in library references.Alternative: code call stack references.The TOP_POSITIVE_LOGITS are `unknown`, `anonymous`, `call`, `calls`. This might mean it's not just any `at <domain>` but one where the context is less identified or involves communication.Looking at `TOKENS_AFTER_MAX_ACTIVATING_TOKEN` again: `org`, `com`, `flash`, `Meteor`. These are domain names and tech names.The texts show `at org...`, `at com...`, `at flash...`, `at Meteor...`.The `TOP_POSITIVE_LOGITS` also has `calls`, `call`, `unknown`, `anonymous`. This could imply a neuron that flags `at` followed by a package name/domain, particularly when it relates to calls or unknown entities in code.However, the prompt asks for a specific pattern. `at org`, `at com`, `at flash`, `at Meteor` is very specific. The `anonymous` and `unknown` are also strong.Let's try to combine: "at" followed by tech domains or "unknown" / "anonymous" contexts.This is getting long.Focus on the most consistent pattern: `at` followed by some identifier.`at org``at com``at flash``at Meteor`This seems to capture the domain names and tech names.What about `anonymous` and `unknown`?The texts contain phrases like "cannot be cast to android.app.DatePickerDialog$OnDateSetListener", "Attempt to invoke virtual method 'android.view.View android.view.Window.findViewById(int)' on a null object reference", "java.lang.NoSuchMethodError". These are error messages where specific calls might be unknown or fail.Let's consider the `TOP_POSITIVE_LOGITS`.`unknown` (2)`unky``anonymous``inka``calls``call``onymous``Unknown``exhausted`The `calls` and `call` are interesting.And `unknown`, `anonymous`.The common pattern in `MAX_ACTIVATING_TOKENS` is `at`.The common pattern after `at` is `org`, `com`, etc. (domain/tech).The positive logits are `unknown`, `anonymous`, `call`.Maybe the neuron detects calls to unknown/anonymous locations, or calls made 'at' these locations."anonymous calls" is short and fits.Let's check if the texts support "anonymous calls".The texts are mostly error logs from Java/Ruby/JS."Attempt to invoke virtual method... on a null object reference" - This could be seen as an 'anonymous call' failing because the object is null."java.lang.NoSuchMethodError" - a call failed because the method is not found.Let's try to connect `at <domain>` with `anonymous` or `call`.The neuron might signal when `at <domain>` is associated with error states or communication.However, the instruction is about finding a pattern in the lists.`at` is the most common MAX_ACTIVATING_TOKEN.`org`, `com`, `flash`, `Meteor`, `` ` `` are after it.The logits `unknown`, `anonymous`, `call` strongly suggest a theme.If we combine `at` with `anonymous` or `unknown`:`at` and `anonymous` or `unknown`.Let's look at the raw data again.`at org.jboss.arquillian...``at System.Threading...``at net.rubygrapefruit...``at org.gradle...``at com.example...``at android.app.ActivityThread``at org.hibernate.hql...``at flash.display...``at BlockDrop_fla::MainTimeline...``at Meteor.absoluteUrl...`The phrase "at <tech domain>" is very strong.The logits `unknown` and `anonymous` are also strong.The phrase must be 3-20 words.Could it be "unknown org/com calls"? That's 4 words.The texts are full of `at org` and `at com`.The logits are `unknown`, `anonymous`, `call`.This seems like a great fit.Let's re-evaluate "at org.jboss.arquillian.protocol.servlet.ServletMethodExecutor.invoke". The neuron activates for "at" and the subsequent path.The logits are for words that *might* follow.The most direct pattern from `MAX_ACTIVATING_TOKENS` and `TOKENS_AFTER_MAX_ACTIVATING_TOKEN` is `at` followed by `org`, `com`, `flash`, `Meteor`.The `TOP_POSITIVE_LOGITS` are words that are *predicted* to be associated with this neuron's activation. This means the neuron *contributes* to predicting these words. So, if the neuron fires strongly, it's more likely to predict `unknown`, `anonymous`, `call`.This implies the neuron detects contexts where `at <domain/path>` is associated with an unknown or anonymous state, or possibly an attempted call.Let's try to capture `at` + `domain/path` + `anonymous`/`unknown`/`call`."at domain names" - too simple, doesn't use logits."anonymous at domain" - good. 3 words."unknown at domain" - good. 3 words."calls at domain" - good. 3 words.Looking at the texts, the `at` is usually part of a stack trace.The `anonymous` and `unknown` relate more to the nature of the error or the object being acted upon.`at org.jboss.arquillian...` is a stack trace line.`at com.example... cannot be cast to...` this is an error, possibly involving an anonymous casting.`Attempt to invoke ... on a null object reference` - implies an unexpected or unknown state.Let's consider the combination of `at <domain>` and `anonymous` or `unknown`.These could be related to abstract or uninitialized components within code execution.How about: "abstract or unknown calls"? 4 words.Or "anonymous calls in code"? 4 words.The prompt states: "Your

    np_acts-logits-general · gemini-2.5-flash-lite
    New Auto-Interp
    Top Features by Cosine Similarity
    Configuration
    decoderesearch/olmo-3-saes/olmo-3-1025-7b
    Prompts (Dashboard)
    24,576 prompts, 128 tokens each
    Dataset (Dashboard)
    monology/pile-uncopyrighted
    No Configuration Found
    Embeds
    IFrame
    Link
    Not in Any Lists

    No Comments

    Negative Logits
    OOD
    -0.08
     grade
    -0.07
    ood
    -0.07
     relative
    -0.07
     sal
    -0.07
     engineers
    -0.07
     nails
    -0.07
     grap
    -0.07
     dri
    -0.07
     blush
    -0.07
    POSITIVE LOGITS
     unknown
    0.11
    unky
    0.11
     anonymous
    0.11
    unknown
    0.10
    inka
    0.10
     calls
    0.10
     call
    0.10
    onymous
    0.09
     Unknown
    0.09
     exhausted
    0.09
    Activations Density 0.315%

    No Known Activations