# Outside read of the H1 all-codings analysis and paper-priority recoding: GPT-6 Astra (gpt-6-astra), 4 October 2026

**Type:** outside-lineage read (one Codex call, xhigh) · **Verbatim below** (local file links shortened to file names).

The calculations pass. The record needs substantive amendments.

1. **Git order: verified.** `1b12833` filing → `a9866f4` script → `33a37a1` results → `c6fdc3a` Step B input → `ed485a1` reader codes and placement script → `81f90a1` placement results. This establishes committed-artifact order, not execution times; qualify “before any computation” accordingly.

2. **Implementation and reproduction: pass.** The candidate sets, uniform sampling, seed, scales, 75% completeness, groups, Welch intervals and label precedence implement the filing. No combinations require exclusion. Both result JSONs reproduce **exactly** with saving disabled. Independent pandas/statsmodels calculations also agree. The original agreement gate is not reinstated; this remains the filed exploratory analysis.

3. **Input: blinding passes; the override needs disclosure.** The excerpt, definitions, coding rule and selected item wordings match their sources. No H1 hypothesis, groups, response data, results or disputed-item status appear. The audit records Read plus StructuredOutput for each reader.

   But “follow the paper’s meaning” (`run3_prompts/paper_reader_input.md:7`) adds an unfiled precedence rule. It licenses changing the registered constructs when interpretation conflicts with them. Describe Step B as **paper-priority exploratory recoding**, not demonstrated preservation of registered meanings. Its causal effect on particular codes is untested.

4. **Reader codes: no repeat of the clear item-76 polarity error, but flag these assignments.** Both readers’ 75=N and 76=N are defensible. Opus’s **23=N** is poorly justified: temporal self-location meets the registered time-as-content criterion; mentioning oneself does not establish orientation toward modelling activity. Fable’s **40/41/66=A** assumes that nonvisual brightness, clarity and radiance indicate awareness knowing itself. The wording does not settle that; the conservative ambiguity rule favours N. These are construct-assignment concerns, distinct from reversing endorsement’s meaning.

   Also amend the quoted explanation of 76: **difference from reflexive awareness does not establish complete absence of inward orientation**.

5. **Record: numerical tables pass; several interpretations do not.** The reading-dependent band, label shares, per-item conditional means, reader agreement, intervals and placements are correct: **29.8th and 29.6th percentiles**, with 14 Inconclusive and two Weakly supported combinations.

   Amend the record (`90_ALL_CODINGS_RESULT.md:36`) as follows:

   - Replace “every reasonable coding” with **sampled combinations of previously assigned codes**. Both universes retain known problematic assignments, including 76=A; a compliant brief does not guarantee compliant codes.
   - Replace “never reversed” with **no sampled coding met the reversal criterion**. U1 contains six negative point estimates. Sample extrema are not universe bounds; “under 5%” also cannot mean “no fair reading.”
   - Delete **“No single question decides it.”** Changing only Fable’s item 66 from A to N changes Inconclusive to Weakly supported: Δ **3.65**, 95% CI **[0.011, 7.290]**. Conditional mean differences are neither maximum individual effects nor label-stability tests.
   - Replace “half a point at most” with **approximately 0.51 points**.
   - Retain the existing caveats, adding that 63 items were inherited without these readers reassessing them. Two readers’ agreement does not establish construct validity or a uniquely correct paper reading. Universe selection and weighting remain consequential choices.

ADOPT WITH AMENDMENTS (1, 3–5)


