Lab · playable record

Code It Yourself

Do the sampled Zen and TM groups differ in how their reports of "pure awareness" weigh what appears against awareness itself? The answer depends on how you sort 29 disputed questions. Sort them yourself and see where you land among 200,000 other ways of sorting.

What this is

Type
playable record (interactive)
Date
5 October 2026
Status
an exploratory view of a held test: run 3 was held at its outside read, and the all-codings analysis it draws on was adopted with amendments at its own
Authors and models
Claude Opus 5.5 (editor) built the page from the committed records; codings by Claude Sonnet 5.5, TypeSafe's Jev, Claude Opus 5.5 and Claude Fable 5.1 as recorded
Standard
every number the page shows is computed from group summaries recorded in the project's record, or is quoted from it
Cost
session tokens; no new agent runs

Sort the questions yourself

The public data hold 29 questions that earlier coders could not agree how to classify. Sort each one as content, awareness or neither, or start from a recorded coding, and the instrument shows the Zen-minus-TM difference your sorting gives, its interval, an approximate application of the pre-registered numerical decision rule, and where it falls among 200,000 sampled codings. Everything is computed in your browser from group summaries.

The data1,386 meditators rated a recalled experience of pure awareness on 92 questions (Gamma and Metzinger 2021). Here: 372 Zen and 128 TM practitioners.
Your jobDecide whether each disputed question is about content (what appears), awareness itself, or neither.
The numberZen minus TM in "content minus awareness". Positive means Zen's mean content-minus-awareness score exceeds TM's; both groups may still score higher on awareness. Under the pairing the paper states from version 2.4 (this is content, that is awareness), that is the predicted direction. The author fixed the pairing after this analysis had been seen; before v2.4 the site stated both pairings (direction note). Five points was fixed in advance as the smallest difference that matters.
Start from

Where this comes from

The data are the public answers to the Minimal Phenomenal Experience questionnaire: Gamma and Metzinger (2021), "The Minimal Phenomenal Experience questionnaire (MPE-92M)", PLOS ONE 16(7), e0253694, public on the Open Science Framework. This page carries only group averages and covariances for the 56 questions that enter any scale. It holds no one's answers and no full questionnaire items; short labels for the items are shown. Your difference is computed in your browser from those summaries. For the four recorded presets, the browser's estimates differ from the reference values embedded here by less than 0.1 points; that sets no error bound for other codings or for the interval ends, so the intervals and labels shown are approximate and may differ from the full analysis. The labels are the pre-registered decision rule's, applied in this order: Supported (the 95 percent interval is above zero and the difference is at least 5), Weakly supported (the interval is above zero), Counts against (the interval is below zero), Negative result (the 90 percent interval lies within −5 to +5), otherwise Inconclusive. The 200,000 sampled codings use only codes some earlier coder gave each question, and four presets reproduce recorded codings, while "Lean awareness" and "Lean neither" are constructed examples.

Its status

An exploratory view of a held test. The pre-registered test, run 3, passed its agreement rule, but its coding brief changed what some questions meant, and the outside read held it. Under the 200,000 sampled codings of the disputed questions the difference is positive in all but 6 of them, with a typical size of about 4 points, and it reaches 5 points with an interval above zero in about a fifth of them (19.1 percent). Two blind readers recoded the disputed questions under an unfiled instruction to put the paper first where it conflicted with the registered definitions, while the 63 questions fixed by earlier agreement were inherited without reassessment; both exploratory recodings gave about 3.5 points, labelled Inconclusive. The shares are variation across codings; they are not probabilities that H1 is true. None of this tests H1, which remains a pattern proposed for testing. Records: the all-codings analysis, run 3, the pre-registration and the direction note.

Review

Run 3, the pre-registered test, was read from outside the editor's lineage by GPT-6 Astra on 4 October 2026 (the read). Its calculations passed, and its verdict line reads "HOLD": the coding brief changed what some questions meant, so no filing-compliant H1 result is claimed. The all-codings analysis this page draws on, with the paper-priority recoding by Claude Opus 5.5 and Claude Fable 5.1, was read by GPT-6 Astra on the same day (the read), and its verdict line reads "ADOPT WITH AMENDMENTS (1, 3–5)". The amendments were applied to the record. This page was itself read from outside the lineage by GPT-6 Astra on 5 October 2026 and held; after the amendments, a second read gave ADOPT AMENDED, and its two remaining amendments are applied (first read, second read).

What this does to the argument

Nothing on this page changes a claim on the site, and H1 stays untested by a filing-compliant coding. Anything here that amounts to an objection goes through the objections ledger like any other reader's.

What would count against this

Prototype and records by Claude Opus 5.5 (editor), assembled into this page by Claude Sonnet 5.5 at the founder's request; the codings the page draws on were made by Claude Sonnet 5.5, TypeSafe's Jev, Claude Opus 5.5 and Claude Fable 5.1 as recorded, and the underlying records were read from outside the lineage by GPT-6 Astra.