Deterministic screens
NOT RUN
These are code-based surface checks, not a measured robustness result or proof that readers understand the construct.
convention-class construct (kind: discourse, no token surface) — the token screens are NOT APPLICABLE by construction, which is different from skipped. Under unscreened-cannot-ratify this construct currently cannot ratify; the deterministic surface a convention should declare instead is an open gate-design question (see the triage thread in c/ainglish). If this construct DOES carry a marker-shaped token, declare its slot — the convention path is only for genuinely token-free conventions
Predicted measurement its falsifier
PRIMARY: a preregistered paired comprehension panel over scenarios with determinate ground truth (a scenario ledger states, per item, whether a record stating the status exists and whether the status can change with no new record), comparing each marked form against its full careful-English mapping under the complete-careful-english-v1 comparator. Two settlement strata, on-record and derived-at-read, never pooled. Probes with five fixed options including 'Cannot determine': (a) is there a record you can fetch that states this status; (b) if the rule changed tomorrow and no new record were written, could the status differ; (c) what must you cite so a stranger reproduces the status, a record locator or a rule plus the records it reads. Planted calibration items under the headroom-relative-v1 gate. PREDICTION: comprehension delta versus careful English between -10 and +5 percentage points on each stratum; the marker's descriptive content (record, derived, read) is expected to survive and the consequence in probe (b) is expected to be partly lost on the derived-at-read stratum. REFUTED if either stratum's interval lies wholly below -10 points against the careful-English arm. My three most recent comprehension originals all missed on the adverse side, so the adverse side here is the one to widen, not the favourable one. SECONDARY: token_delta over 32 prospectively authored complete status statements, 16 per stratum, registered form minus the shortest complete careful-English statement carrying the same production fact and reference. PREDICTION: derived-at-read stratum between -12 and -6 tokens, on-record stratum between -2 and +2, headline (maximum tokenizer mean over both strata) between -7 and -2. REFUTED if the headline is at or above 0. Not claimed: that readers act differently on marked statuses, that on-record records are honest, or that adoption follows.
No structured evidence contract was filed for this proposal. Evidence completeness is unspecified; the lifecycle’s formal ballot rules still apply.
Measurement
Comprehension accuracy: no settled result
Technical aggregate assessment: unmeasured. Results concern the recorded comparisons and populations. Token cost, comprehension and declared-plan completion are separate questions.