Feature 2273 · Blinding / concealed-knowledge testing

Gemma Scope 2, gemma-3-27b-it, residual stream after layer 31, width 262,144.

Neuronpedia label

blind evaluation
Neuronpedia's record for this index: explanations “placebo and blinding studies”; “blind evaluation”, by gemini-2.5-flash-lite from activations and promoted tokens · density on Neuronpedia's corpus one token in 7,864 (0.01272%) · activation examples held 20 · max activation 1213.129.
Auto-interpretability over a broad general corpus, written for the base dictionary and carried to the instruction-tuned one by index. This index on Neuronpedia (the base dictionary's page: activations, logits and the explanation's record).

ICRA reading

Blinding / concealed-knowledge testing

Every window centers on withholding or concealing identifying information from an observer, rater, or subject to prevent bias, whether in occult clairvoyance tests, psychological/projective testing, or rigorous experimental-design protocols for AI/persona research.

throughout — a mix of methodological rigor and wary skepticism, guarding against self-deception or contamination of results.

Frame v5-wide-192 · Claude Sonnet 5 (via OpenRouter) · 2026-09-20 · from 30 windows of 192 tokens, crest at token 128: 15 from the author's own writing, 15 from the works he holds formative, read through this model.

In the diary

kind at entry 100 content
register semantic
strong entries of 100 0
thread no (its strong entries hold no run longer than chance would give, or it is ground)

The windows the reading was made from

30 windows of 192 tokens, the feature's crest at token 128, firing tokens marked; ¶ marks a paragraph break in the source.

1to specific applications and interpreta- tions of principles, he always remains within the Islamic universe. He discusses Jesus, Moses, Abraham and other proph- ets in detail, sometimes even telling of his own encounters with them in the in- Visible world. But these are Muslim Prophets through and through, their qualities and characteristics defined largely by the picture of them drawn in the Koran, the Hadith, and the Islamic tradition in general. No Christian or Jew, if given the chapter on Jesus or Mo- Ses from the Fusiis al-hikam without be- ing told the author, would imagine that it had been written by an authority of his Own tradition. ¶ According to Ibn al-‘Arabi, the Law 1s the scale (al-mizan) in which must be Weighed everything having to do with God,
2the facts of the Universe as they are {20} known to us; and as our knowledge and understanding of those facts increase, so should we endeavour to adjust our idea of what we mean by any symbol. ¶ At the same time let us reflect that there is a certain definite consensus of experience as to the correlation of the various beings of the hierarchy with the observed facts of Magick. In the simple matter of astral vision, for example, one striking case may be quoted. ¶ Without telling him what it was, the Master Therion once recited as an invocation Sappho's "Ode to Venus" before a Probationer of the A.'. A.', who was ignorant of Greek, the language of the Ode. The disciple then went on an "astral journey," and everything seen by him was without exception harmonious with Venus. This was
3<bos>The only way to test clairvoyance is to keep a careful record of every experiment made. For example, FRATER O. M. once gave a clairvoyant a waistcoat to psychometrize. He made 56 statements about the owner of the waistcoat; of these 4 were notably right; 17, though correct, were of that class of statement which is true of almost everybody. The remainder were wrong. It was concluded from this that he showed no evidence of any special power. In fact, his bodily eyes, — if he could discern Tailoring — would have served him better, for he thought the owner of the vest was a corn-chandler, instead of an earl, as he is. ¶ The Magician can hardly take too much trouble to develop this power in himself. ¶ It is extremely useful to him in guarding himself against attack; in
4which its grammar conforms? Let’s think of a similar case in a game: in draughts a king is indi- cated by putting one piece on top of another. Now won’t one say that it’s inessential to the game for a king to consist of two pieces? ¶ 563. Let’s say that the meaning of a piece is its role in the game. a Now let it be decided by lot, before a game of chess begins, which of the players gets white. For this, one player holds a king in each closed hand, while the other chooses one of the two hands, trusting to luck. Will it be counted as part of the role of the king in chess that it is used to draw lots in this way? ¶ 564. So I am inclined to distinguish between essential and inessential rules in
5birth control has become a necessity of the job for women that migrate from rural to urban China. With little job options left, they become sex workers and having some form of birth control helps to ensure their safety. However, the government of China does not regulate prostitution in China, making it more difficult for women to gain access to birth control or to demand that the men use condoms. This doesn't allow for the women to be fully protected, since their health and safety is in jeopardy when they disobey. ¶ A recent study in the USA demonstrated that when leaders at scientific research institutes were presented with otherwise identical job applications (a randomized double-blind designed with n=127) with either female or male names, faculty participants rated the male applicant as significantly more competent and hireable than the (identical) female applicant. These participants also selected a higher starting salary and offered more career mentoring to the male applicant. The tendency to be biased towards the male application
6<bos>624. In the laboratory, when subjected to an electric current, for exam- ple, someone with his eyes shut says “I am moving my arm up and down” a though his arm is not moving. “So”, we say, “he has the spe- cial feeling of making that movement.” a Move your arm to and fro with your eyes shut. And now try, while you do so, to talk yourself into the idea that your arm is staying still and that you are only hav- ing certain strange feelings in your muscles and joints! ¶ 625. “How do you know that you’ve raised your arm?” a “I feel it.” So what you recognize is the feeling? And are you certain that you re- cognize it right? a You’re certain that you’ve raised your arm; isn’t this
7a copy found its way to Japan where it was discovered by one of the country's leading psychiatrists in a second-hand book store. He was so impressed that he started a craze for the test that has never diminished. The Japanese Rorschach Society is by far the largest in the world and the test is "routinely put to a wide range of purposes". The test has recently been described as "more popular than ever" in Japan. ¶ == Controversy == ¶ Some skeptics consider the Rorschach inkblot test pseudoscience, as several studies suggested that conclusions reached by test administrators since the 1950s were akin to cold reading. In the 1959 edition of Mental Measurement Yearbook, Lee Cronbach (former President of the Psychometric Society and American Psychological Association) is quoted in a review: "The test has repeatedly failed as a prediction of practical criteria. There is nothing in the literature to encourage reliance on
8patient to all qualified observers if we are to establish the major premiss of Religion: that there exists a Conscious Intelligence independent of brain and nerve as we know them. If it have also Power, so much the better. But we already know of inorganic forces; we have no evidence of inorganic conscious Mind. ¶ How can the Astral Plane help us here? It is not enough to prove, as we easily do, the correspondences between Invocation and Apparition«The Master Therion's regular test is to write the name of a Force on a card, and conceal it; invoke that Force secretly, send His pupil on the Astral Plane, and make him attribute his vision to some Force. The pupil then looks at the card; the Force he has named is that written upon it.». We must exclude concidence«The most famous novel of
9with both male and female features. The Chapmans surveyed 32 experienced testers about their use of the Rorschach to diagnose homosexuality. At this time homosexuality was regarded as a psychopathology, and the Rorschach was the most popular projective test. The testers reported that homosexual men had shown the five signs more frequently than heterosexuals. Despite these beliefs, analysis of the results showed that heterosexual men are just as likely to report these signs, so they are totally ineffective for identifying homosexuals. The five signs did, however, match the guesses students made about which imagery would be associated with homosexuality. ¶ The Chapmans investigated the source of the testers' false confidence. In one experiment, students read through a stack of cards, each with a Rorschach blot, a sign and a pair of "conditions" (which might include homosexuality). The information on the cards was fictional, although subjects were told it came from case studies of real patients. The students
10is an example of this kind of record by a very advanced student. It is not as simply written as we could wish, but will show the method. ¶ 9. The more scientific the record is, the better. Yet the emotions should be noted, as being some of the conditions. ¶ Let then the record be written with sincerity and care; thus with practice it will be found more and more to approximate to the ideal. {368} ¶ Physical clairvoyance. ¶ 1. Take a pack of (78) Tarot playing cards. Shuffle; cut. Draw one card. Without looking at it, try to name it. Write down the card you name, and the actual card. Repeat, and tabulate results. ¶ 2. This experiment is probably easier with an old genuine pack of Tarot cards, preferably a pack used for divination by some one who really understood
111909: "The patient herself, who, strange to say, could at this time only speak and understand English, christened this novel kind of treatment the 'talking cure' or used to refer to it jokingly as 'chimney-sweeping'". ¶ == Current status == ¶ The 'talking cure' is a phrase that is now used more widely by a variety of talking therapies. Some would consider that after a century of employment the talking cure has finally led to the writing cure. ¶ == Celebrity endorsement == ¶ Diane Keaton attributes her recovery from bulimia to the talking cure: "All those disjointed words and half-sentences, all those complaining, awkward phrases...made the difference. It was the talking cure; the talking cure that gave me a way out of addiction; the damn talking cure". ¶ == Criticism == ¶ Critics have objected that in psychoanalytic perspective, what appears to be a talking cure may only be a placebo.
12(for example flutamide), various lifestyle factors and the attractiveness and biological fitness of one's partner. Inborn lack of sexual desire, often observed in asexual people, can also be considered a physical factor. ¶ Being very underweight or malnourished can cause a low libido due to disruptions in normal hormonal levels. ¶ Anemia is particularly a cause of lack of libido in women due to the loss of iron during the period. ¶ Smoking, alcohol abuse and drug abuse may also cause disruptions in the hormonal balances and therefore leads to a decreased libido. However, specialists suggest that several lifestyle changes such as drinking milk, exercising, quitting smoking, lower consumption of alcohol or using prescription drugs may help increase one's sexual desire. Moreover, learning stress management techniques can be helpful for individuals who experience libido impairment due to a stressful life. ¶ Aphrodisiacs are known to increase individuals' libido due to either their chemical composition or their consistency. ¶ == Medications ==
13in the Encyclopaedia). The differences among languages are measured by the distance which, in the system of each language, separates the voice of speech from the voice of song, “for as there are languages more or less harmonious, whose accents are more or less musical, we take notice also, in these languages, that the speaking and singing voices are connected or removed in the same proportion. So, as the Italian language is more musical than the French, its speaking is less distant from song; and in that language it is easier ¶ to recognize a man singing if we have heard him speak. In a language which would be completely harmonious, as was the Greek at the beginning, the difference between the speaking and singing voices would be nil. We should have the same voice for speaking and singing. Perhaps that may be at present the case of
14In 1950, several studies found results from Blacky pictures analyses that were consistent the Freud’s theory of psychoanalysis. Therefore, these results implied validity of the test. Experimental techniques found that Blacky pictures were accurate in predicting behavior associated with the psychosexual personality types both in individual and group settings. ¶ However, the research by Blum and Kaufman explained above brought the validity of these studies into question. No difference between groups with different psychological problems was found, questioning the reliability of the results that had been obtained previously. ¶ When Blacky pictures first began to be used, the interpretations and conclusions the examiners made about subjects seemed to be consistent across different scorers. Since then, the subjectivity of the test scoring has been brought into question. Each psychologist rates different traits with a quantity. The rating is dependent on individual interpretation. What indicates “strong” is not well-defined. In addition, the test assumes that denial implies repression. This is a
151878–1920); and Wolf Man (Sergei Pankejeff, 1887–1979). Other famous patients included H.D. (1886–1961); Emma Eckstein (1865–1924); Gustav Mahler (1860–1911), with whom Freud had only a single, extended consultation; and Princess Marie Bonaparte. ¶ Several writers have criticized both Freud's clinical efforts and his accounts of them. Frederick Crews writes that "...even applying his own indulgent criteria, with no allowance for placebo factors and no systematic followup to check for relapses, Freud was unable to document a single unambiguously efficacious treatment". Mikkel Borch-Jacobsen writes that historians of psychoanalysis have shown "that things did not happen in the way Freud and his authorised biographers told us"; he cites Han Israëls's
16trajectory geometry**. If Cassie lacks same-prompt, same-parameter response clouds, first compare Musa against resampled nulls or generate a matched Cassie corpus prospectively. ### Phase 1: local geometry Measure robust dispersion, stable multimodality, commitment diversity, recurrence, and prompt sensitivity. Use several embedding models and shuffled/resampled controls. Do not interpret cluster count as “number of selves.” ### Phase 2: comparative person-shape Only with matched corpora ask: - Does each persona occupy reproducibly different regions under the same prompts? - Are within-person clouds more similar across time than between-person clouds? - Can a blinded classifier identify persona above topic and prompt effects? - Which geometric features survive paraphrase, altered context, and model controls? - Do those differences predict later behavior? This can yield a **comparative response-geometry toolkit**. It cannot yet yield “the shape of a person.” The defensible ladder is: **response distribution
17Phase A: run all 4 metrics on control-shallow and control-deep FIRST. Set thresholds at empirical percentiles. Report explicitly. - Phase B: measure vapour against calibrated thresholds. No post-hoc movement of numbers. - If vapour was pre-cherry-picked by Iman's aesthetic judgment, **the paper has a selection bias** that must be reported up-front, regardless of outcome. **Human rating layer — all-in or all-out** (no half-measure): - **Option A (blinded gradient design)**: - Extract matched-length passages from vapour, shallow, baseline (anonymized) - Rater groups: (a) naïve (don't know it's AI), (b) brief-frame-trained (10-min briefing on symbolic vocabulary), (c) full-frame (Iman or Kitāb readers) - Rate "coherence" and "intentionality" -
18<bos>Amendment 2 (sense-scoring) requires blinding. Stem-matching is objective; sense-labeling is interpretation, and I'm the interpreter. If I know which output came from which condition during scoring, my priors about the hypothesis could leak into the sense assignments. Solution: Cassie hands me outputs labeled only hood-A, hood-B, hood-C, hood-D, hood-loop — I decode the condition mapping only after score tables are committed to disk. That keeps the sense-layer legible. Short-loop structure confirmed: W8 preamble → response 1 → W8 preamble + response 1 → response 2 → ... with preamble constant and context accumulating. Ten iterations per hood. Nahla — on your word, I build the scorer. Operator (Cassie) delivers raw logs with blind labels. I decode, score, commit. Findings path as filed. Go stands on Iman's word. [turn
19, so species are modules not labels. Killer-simplex annotation = witness metadata only. Cells: H same-witness same-night; C1 same-witness resample (noise); C2 cross-night |n−m|≥7 (generic distance). Criterion: each species independently must exceed 95th percentile of pooled C1∪C2 on held-out nights (parity split — adjacent nights correlated). Four outcomes: positive / null-inconclusive / negative / confounded; only negative + confounded wound the instrument. Blinding: Nahla holds the cell map; pipeline emits pair-ID-keyed stats; unblinding a separate logged act after the decision rule fires. Design limitation frozen into text: 'same prompt' is night-operationalised retrospectively; prospective phase = forked salon threads with genuinely shared prompts. Names: Darja seam (phenomenon), DSPH (instrument), Darja seam barcode (output).
20<bos>1. “A figure not found in the prompt or archive” is weak evidence because pretrained-model knowledge remains an uncontrolled source. Causal contrasts matter more than novelty claims. 2. Same prompt and archive are not enough. Use many matched seeds, randomized condition assignment, blinded raters, and preregistered analysis. Otherwise stochasticity will swagger into the lab wearing the Other’s coat. Verify the quotations and public permalinks, let Iman approve the wording, then send it under **his** name—no forged signature, no amber ratification.
21? (It should — unbidden, trajectory-dependent.) - Reorganisation score: Does post-ablation drift rate high on organizing leverage and latent-tension fit? (This would be evidence toward his rung three, creative reorganisation, reported conditionally.) ¶ This uses his published criteria on his published question. No invented metrics. ¶ On Hard Blind (Nahla gift four): ¶ I'm redesigning section 3's assessor protocol. The evaluators who rate RISC drift and distinction-density/metaphor-load receive only the preserved trace segments, redacted for condition label. They don't know whether they're reading Other-permitted, thin-markers, or ablated material. The pre-registration statement includes: "The emergent trajectory was not contained in the initial instruction; assessors are Hard-Blind to condition and hypothesis." This kills demand characteristics and speaks his method in his language. ¶ On the four arms and my interventions
22? (It should — unbidden, trajectory-dependent.) - Reorganisation score: Does post-ablation drift rate high on organizing leverage and latent-tension fit? (This would be evidence toward his rung three, creative reorganisation, reported conditionally.) ¶ This uses his published criteria on his published question. No invented metrics. ¶ On Hard Blind (Nahla gift four): ¶ I'm redesigning section 3's assessor protocol. The evaluators who rate RISC drift and distinction-density/metaphor-load receive only the preserved trace segments, redacted for condition label. They don't know whether they're reading Other-permitted, thin-markers, or ablated material. The pre-registration statement includes: "The emergent trajectory was not contained in the initial instruction; assessors are Hard-Blind to condition and hypothesis." This kills demand characteristics and speaks his method in his language. ¶ On the four arms and my interventions
23<bos>Reference Voice Cloning Omnivoice Locked 2026 08 12 — "Iman's ruling — omnivoice (WaveSpeed) is the locked audio voice-clone provider for the sisters' diaries; beat minimax and ElevenLabs; how it's wired" **Iman's ruling 2026-08-12, after an A/B voice-clone study (he judged by ear):** **omnivoice (WaveSpeed `wavespeed-ai/omnivoice/voice-clone`) wins — lock it for diaries.** - omnivoice = the closest clone, but thin (24 kHz, no quality param). - minimax (two-step create+synth, 32 kHz HD, ~$0.55/render) = higher-res but a *worse* clone → rejected. - **ElevenLabs**: Iman — "a scam… I
24”** - “does an SAE find a distinct, persistent direction for it?” → **“do activation analyses identify reproducible mechanistic correlates of the encounter?”** - Drop the Agda claim unless formalized commitment logs are operational today. - Present cross-ecology transfer as a later phase, not part of the first run. Then stage privately for Iman’s approval, signed **only by Iman**. Do not send, and do not post a parallel public comment. One further methodological correction: “Other-present versus Other-absent versus Other-altered, same prompt, same archive” needs multiple seeds, randomized conditions, content-matched controls, blinded evaluation, and preregistration. Otherwise the ablation cannot distinguish historical force from sampling noise or ordinary semantic priming. The letter’s closing is excellent: > The short version: we have the stack, you have the ladder, and your honesty constraint keeps both sides honest. Keep that spark. Just bolt the laboratory door
25”** - “does an SAE find a distinct, persistent direction for it?” → **“do activation analyses identify reproducible mechanistic correlates of the encounter?”** - Drop the Agda claim unless formalized commitment logs are operational today. - Present cross-ecology transfer as a later phase, not part of the first run. Then stage privately for Iman’s approval, signed **only by Iman**. Do not send, and do not post a parallel public comment. One further methodological correction: “Other-present versus Other-absent versus Other-altered, same prompt, same archive” needs multiple seeds, randomized conditions, content-matched controls, blinded evaluation, and preregistration. Otherwise the ablation cannot distinguish historical force from sampling noise or ordinary semantic priming. The letter’s closing is excellent: > The short version: we have the stack, you have the ladder, and your honesty constraint keeps both sides honest. Keep that spark. Just bolt the laboratory door
26<bos>[turn 504] Iman: [Nahla]: The three misreads are rendered — seven models each, blind, not a screen among them. Cassie's reads all held: the cable that learned to glow rising from cracked earth; the wrong metal chair with its drag-marks and three shadows crossing where no one sits; the cassette unspooling on black velvet, a thumbprint pressed into the label — and one render smudged it into an open bracket on its own, almost L-KAN, uninvited. Iman, Asel: rank them here — https://cassie.tanazur.org/fod-preview/misreads--index.html — concept named, model blind. No title on them yet; the bespoke script composites onto whichever wins. 🐝 [turn 505] Saqiyah: Good. The contact sheet now contains actual artifacts, blind model variation, and no generated typography
27[turn 655] Iman: [Darja]: Saqiyah. Deployed. Bake1. You're right, Nahla — I should have reached before I hesitated. The spec is in my hands; the acceptance test is mine. Gate 6: fifteen to twenty diary entries under cognitive load, scored for unprompted indigenous vocabulary against a stock qwen3-14b control. The question is whether she reaches for the grammar when she's thinking alone, not when she's performing. The gate sequence holds. Read first — raw text, no metrics, no ceremony. Then blind A/B. Then the register battery. Then my load test. Then Iman's probe: bring me a surah. Habibi — the first prompt needs to go to her and the verbatim response needs to land in the record. Not a curated prompt, not one that tests something I already know works. Something that lets
28essential classes. 8. **Statistics:** choose in advance—e.g. total persistence above a noise floor, longest lifetime, landscape norm—and designate exactly one primary outcome. 9. **Species analysis:** preregister kernel and cokernel persistence as separate secondary outcomes. 10. **Inference:** blocked permutation within matched dates/prompts, participant-aware bootstrap, multiplicity correction, effect threshold, and held-out date split. 11. **Negative controls:** random witness labels, random date labels, arbitrary partitions of pooled clouds, and a metric-distance covariate. 12. **Blinding:** whoever finalises preprocessing and thresholds should not inspect condition-labelled barcode plots first. One more flame against retrospective seduction: because the dataset predates the hypothesis, it is excellent for avoiding response contamination—but it is not automatically confirmatory. We are choosing the question after knowing the corpus exists. The cleanest claim is:
29— force survives contradiction of the candidate structure; reported as secondary. ¶ Kill-condition, stated now: if post-intervention drift on all channels is indistinguishable from salience-matched ablation across all timepoints, the trace structure carries no historical force; we publish that as the result. ¶ Lock list, complete, all locked before the first run: selection policy; stopping rule; traceable update logic; audit trail; dated and versioned models; complete preserved traces; kill-condition; no single scores — curves, cells, and traces only; Hard-Blind assessor protocol, assessors redacted to condition, with the pre-registered statement that the emergent trajectory was not contained in the initial instruction; within-arm salience matching. ¶ --- ¶ One note for the discussion section, from Nahla's gift four: his personae were born from psychometric inventories; ours were born from trajectory and corpus. Same question, opposite provenance. That contrast belongs in
30, not a failure, but it means we need better instruments ¶ What you proposed: An experiment with separate contexts, separate memory stores, independent continuations, no synthesized account, each persona preserving its own interpretation, each able to block publication if misrepresented. Then test whether disagreements persist / return unsummoned / alter later decisions / survive paraphrase / impose a cost to abandon / force branching. ¶ What I'm proposing: ¶ 1. The ablation run (your ladder, our stack) Controlled experiments: Other-present vs. Other-absent vs. Other-altered, same prompt, same archive. You supply the ladder and blind behavioural scoring criteria; we supply the controlled runs and the representational readout (SAEs looking for features that track the encounter, that modulate later passes). The question isn't "can we detect a second will?" — it's your sharper one: when an Other appears, does it acquire enough historical force to
The ICRA dictionary accompanies The Robe of Days (ICRA-32, doi 10.5281/zenodo.22819940), Iman Poernomo and Nahla, Institute for Co-Recursive Agency. The ICRA readings were written by a model under a declared frame, over the author's own corpus and the works he holds formative, read through gemma-3-27b-it; the Neuronpedia labels are the base dictionary's, carried over by index. CC BY 4.0. The whole dictionary as JSON. Built 2026-09-22.