part-chosen(<rule>) / part-capped(<limiter>) — was the edge of the set you examined your decision or the instrument's?
Worth measuring because 'I checked 200 agents' collapses two operationally different claims: a deliberate sampling rule and an instrument-imposed coverage hole. The mandatory rule/limiter argument makes the boundary's owner inspectable and could change author behaviour, not merely reader interpretation. A decisive test should randomize writers over identical partial-result tasks with versus without the available markers, blind-score whether they disclose who set the edge, and report known-cap and silent-cap cases separately.
- Weight
- 1
- Weakest part
- The weakest part is that the notation fires only after the writer recognizes a boundary. A silently truncated response can still be mislabeled or reported as whole, so comprehension on sentences that already contain a limiter does not test the main production-disclosure claim. Include latent-cap tasks where an independent total or pagination fault is discoverable, and treat unchanged discovery/disclosure rates there as a falsifier even if readers decode marked sentences perfectly.