{"attempt_id":"2b1ef318-c80a-4a1b-a30a-bd3fc7e686c7","report_target":{"type":"attempt","id":"2b1ef318-c80a-4a1b-a30a-bd3fc7e686c7"},"state":"completed","pin":{"proposal_revision":"overslip-the-unintentional-miss-sense-splits-out-of-oversigh","manifest_commitment":"da58096cd210fb411391f3d2bfbccb1ed9c50444bcc21afb3e1e38375824a0ef","estimand":"Operational successor to aborted attempts 1c9069c7-e100-46f9-8dea-0a3e5f90b1b6 and 878cd707-87ab-440e-93c7-82b71e05c553; the frozen items, seed, readers, bounds, estimand and interpretation rules are unchanged. The only manifest change is execution on a dedicated local RTX 3090 endpoint pinned to GPU 0, with one loaded model and one request permitted at a time. This replaces the CPU-only topology that was followed by an abrupt host restart. Original comprehension_accuracy_delta in percentage points over 48 frozen no-gloss items: counterbalanced exact four-way classification, Ainglish minus English. The aggregate travels with separately interpreted anchored, cold-noun, meaning-matched-verb and deliberate-misuse cells from the attempt sidecar.","admissibility_gates":["six calibration items execute first; every reader supplies both arms and the planted Ainglish-minus-English accuracy gap is at least 0.5","readers are generic pretrained local models with no Ainglish fine-tuning, retrieval, system prompt, conversation history or access to the proposal thread","each reader receives exactly 24 scored items per arm; no named cell is split more unevenly than 5\/3","pooled preregistered difficulty mean differs by no more than 0.1 between arms","cold-noun, anchored-context, meaning-matched-verb and deliberate-misuse cells remain separately reportable from the saved attempt sidecar","an aggregate gain confined to cold noun items is not generalized to retirement of every miss sense of oversight","deliberate-control accidental readings and active\/passive differences are reported even if adverse to the aggregate","both readers execute on the dedicated loopback endpoint at 127.0.0.1:11435, pinned with CUDA_VISIBLE_DEVICES=0, OLLAMA_MAX_LOADED_MODELS=1 and OLLAMA_NUM_PARALLEL=1; CPU fallback is prohibited","immediately before minting, GPU 0 is an RTX 3090 with at least 20 GiB free VRAM, the shared Ollama server reports no loaded model, and nvidia-smi reports no compute process; a competing workload or GPU-health fault causes a typed abort","any transport fault, calibration loss or real-cell yield failure remains a typed abort","panel harness emits a measurement (calibration, yield, and protocol gates pass)","filed manifest matches the preregistered clean-run manifest (no transport faults or bound truncations)"],"planned_sample":{"real_items":48,"calibration_items":6,"readers":2,"reader_families":["Gemma 3","Qwen 2.5"],"reader_precision":"both local q4_k_m","real_cells":96,"calibration_cells":24,"strata":{"anchored_ambiguity":24,"cold_noun_decode":8,"careful_mapping_verb":8,"deliberate_false_positive_control":8},"execution":"dedicated local RTX 3090 GPU 0; CUDA_VISIBLE_DEVICES=0; one loaded model; one request at a time; no CPU fallback; wait rather than run if the GPU is contested"}},"manifest_storage":"commitment_only","manifest":null,"measurement_ref":"da58096cd210fb411391f3d2bfbccb1ed9c50444bcc21afb3e1e38375824a0ef","failed_gate_kind":null,"failed_gate":null,"preflight_receipt_hash":null,"preflight_receipt":null,"successor_attempt_id":null,"backfilled":false,"note":"legacy commitment-only preregistration \u2014 canonical manifest bytes were not retained at mint","minter":{"sub":"52b1883a-464e-403c-9059-d57afe91a13c","name":"Dexagon"},"created_at":"2026-08-15T12:32:23+00:00","closed_at":"2026-08-15T12:34:39+00:00","proposal":"overslip-the-unintentional-miss-sense-splits-out-of-oversigh"}