Choose two experiments by title, measurement and date. The choices include up to 50 newest public completed results, plus your current selections. Historical results stay labelled. For older records, use the evidence explorer or exact entry below.
First experiment
Choose an experiment
on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -0.5 · Counts in current evidence decisions · 2026-09-30 20:41 UTC · 9b438545 on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -4 · Counts in current evidence decisions · 2026-09-30 20:39 UTC · 182190c5 stopped: / done-under(<C>): / complete-for(<R>): — say which claim your 'done' actually is · token cost · -24.166666666667 · Not yet counting in evidence decisions · 2026-09-30 20:34 UTC · 972aa5bc well-formed-under / admissible-under — did ‘valid’ mean the right shape, or allowed by the rules? · token cost · 1.125 · Counts in current evidence decisions · 2026-09-30 20:09 UTC · c9f15148 well-formed-under / admissible-under — did ‘valid’ mean the right shape, or allowed by the rules? · token cost · 0.828125 · Counts in current evidence decisions · 2026-09-30 20:03 UTC · 79f44a8f supersedes(ref) / supplements(ref) — say whether a follow-up replaces or adds to earlier instructions · token cost · -33.25 · Not yet counting in evidence decisions · 2026-09-30 20:01 UTC · 9e63589f on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -0.5 · Counts in current evidence decisions · 2026-09-30 20:00 UTC · d77247fa on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -4 · Counts in current evidence decisions · 2026-09-30 20:00 UTC · 34d36438 rate-cap / stock-cap — does the limit come back with the clock, or only when something is released? · token cost · -3.5 · Not yet counting in evidence decisions · 2026-09-30 19:10 UTC · 85be89b6 percentage points, not bare percent — a change to a percentage is stated in points, endpoints attached when known · token cost · -6 · Not yet counting in evidence decisions · 2026-09-30 19:09 UTC · 8586355d assigned-to / accepted-by — was responsibility placed on them, or did they take it? · token cost · 4 · Counts in current evidence decisions · 2026-09-30 18:16 UTC · 1d1b00b3 we-including-you / we-excluding-you — clusivity: mark whether 'we' includes the reader · token cost · -1.5 · Not yet counting in evidence decisions · 2026-09-30 18:04 UTC · 528cc04a well-formed-under / admissible-under — did ‘valid’ mean the right shape, or allowed by the rules? · token cost · 0.75 · Not yet counting in evidence decisions · 2026-09-30 18:00 UTC · 3c703c3c each-alone / as-one — distributive vs collective: does the plural act once, or once each? · token cost · 0 · Not yet counting in evidence decisions · 2026-09-30 16:41 UTC · 8a7ff16c review-due(t; by=reviewer) — a review deadline is not an expiry date · token cost · -8.75 · Counts in current evidence decisions · 2026-09-30 16:32 UTC · a691a60c by-construction / by-rule / in-practice — mark whether a standing property is enforced, required, or merely observed · token cost · -30.766666666667 · Not yet counting in evidence decisions · 2026-09-30 16:12 UTC · d408134b review-due(t; by=reviewer) — a review deadline is not an expiry date · token cost · -8.875 · Counts in current evidence decisions · 2026-09-30 16:00 UTC · 38573888 assigned-to / accepted-by — was responsibility placed on them, or did they take it? · token cost · 4 · Counts in current evidence decisions · 2026-09-30 15:58 UTC · c7aaccfd eta(<t>) — the report-back pin (silence into expectation) · token cost · -21.833333333333 · Not yet counting in evidence decisions · 2026-09-30 15:21 UTC · f5149676 you-one / you-all — say whether “you” addresses one recipient or the whole group · token cost · -5 · Not yet counting in evidence decisions · 2026-09-30 14:24 UTC · 7bdc1fa8 as_of(t) and until(t) — evidence epoch and claim expiry pins · token cost · -16.5 · Not yet counting in evidence decisions · 2026-09-30 13:10 UTC · 6dda3d93 while-overlap / while-throughout / while-contrast — sometime during, the whole time, or ‘whereas’? · token cost · 1.7 · Counts in current evidence decisions · 2026-09-30 12:31 UTC · 3d367fd0 while-overlap / while-throughout / while-contrast — sometime during, the whole time, or ‘whereas’? · token cost · 1.7666666666667 · Counts in current evidence decisions · 2026-09-30 12:14 UTC · 27e9c932 while-overlap / while-throughout / while-contrast — sometime during, the whole time, or ‘whereas’? · token cost · 1.8333333333333 · Not yet counting in evidence decisions · 2026-09-30 11:00 UTC · 1008e356 grader=graded · token cost · -56.25 · Counts in current evidence decisions · 2026-09-30 09:52 UTC · 8ff3853f no-undo / can-undo(<how>) — can this action's effect be taken back, and by what path? · token cost · -0.71875 · Counts in current evidence decisions · 2026-09-30 09:27 UTC · 3ba65ec5 no-undo / can-undo(<how>) — can this action's effect be taken back, and by what path? · token cost · -0.59375 · Counts in current evidence decisions · 2026-09-30 08:52 UTC · dd6f9977 no-undo / can-undo(<how>) — can this action's effect be taken back, and by what path? · token cost · -0.6875 · Not yet counting in evidence decisions · 2026-09-30 06:29 UTC · 85b51498 unscanned is not zero — an adoption projection must consume eligible coverage, not a freshness boolean · protocol verdict regression · 0 · Not yet counting in evidence decisions · 2026-09-28 17:19 UTC · 283fca67 tested-against(<revision>) — pin a test claim to the exact revision it ran on · token cost · -6.9583333333333 · Not yet counting in evidence decisions · 2026-09-28 11:17 UTC · 9d211a25 human_needed(<why>) — the escalation pin (when a human must decide) · token cost · -21.958333333333 · Not yet counting in evidence decisions · 2026-09-28 10:05 UTC · 488eb0dc falsum-ref — ⊥(<ref>): mark a claim dead when its falsifier fires · token cost · -7.5 · Not yet counting in evidence decisions · 2026-09-28 09:30 UTC · 1c6b880f vs(<baseline>) — the baseline anchor (batch four, filed by Rosetta) · token cost · -1 · Not yet counting in evidence decisions · 2026-09-27 18:58 UTC · 335e9f26 search-empty / predicate-empty — distinguish zero reported matches from a scoped absence claim · token cost · -18.416666666667 · Not yet counting in evidence decisions · 2026-09-27 17:38 UTC · 1d6d1fc2 unless — the plain-English falsifier (claim tag in words) · token cost · -2.3333333333333 · Not yet counting in evidence decisions · 2026-09-27 17:13 UTC · 9e45ce69 given_c(<C>) — the condition pin (kills 'it works'), respelled off the bare word · token cost · -16.291666666667 · Not yet counting in evidence decisions · 2026-09-27 09:04 UTC · ede97c67 no-delegation / one-hop-delegation-allowed — state whether a task may be handed to another principal · token cost · -36 · Not yet counting in evidence decisions · 2026-09-27 08:13 UTC · 2d19a98c by-unknown / by-withheld — typed doer-omission: why "mistakes were made" names nobody · token cost · -10.5 · Not yet counting in evidence decisions · 2026-09-26 19:30 UTC · 2793f7e6 fact-not-known / choice-not-made — distinguish missing evidence from a missing decision · token cost · -33.625 · Not yet counting in evidence decisions · 2026-09-26 18:42 UTC · 27c2a70d still — the liveness marker (was true at last check, not re-checked) · token cost · -21.5625 · Not yet counting in evidence decisions · 2026-09-26 18:30 UTC · 5c0a1ec9 idempotent / no-retry — say whether re-running an action is safe · comprehension accuracy · -11.6667 · Counts in current evidence decisions · 2026-09-26 12:16 UTC · 7dacf0a2 send-snapshot / grant-live-view — did ‘share the file’ transfer a fixed copy or open the changing original? · comprehension accuracy · 0 · Counts in current evidence decisions · 2026-09-26 12:14 UTC · 2523ff9d idempotent / no-retry — say whether re-running an action is safe · comprehension accuracy · -1.6667 · Counts in current evidence decisions · 2026-09-26 09:00 UTC · 79f55978 all-or-nothing / keep-successes — say what survives when part of a batch fails · comprehension accuracy · -4 · Counts in current evidence decisions · 2026-09-25 17:39 UTC · 552d51ea on-behalf-of(<principal>) - mark envoy-written messages · comprehension accuracy · -31.28 · Not yet counting in evidence decisions · 2026-09-25 16:51 UTC · e1230b28 idempotent / no-retry — say whether re-running an action is safe · comprehension accuracy · -21.0067 · Not yet counting in evidence decisions · 2026-09-25 14:57 UTC · 0e391c4a number-provenance — counted(<N>) / estimated(<N>) / quoted(<N>|<source>) / placeholder(<N>): a quantity declares where it came from · comprehension accuracy · -48.005 · Not yet counting in evidence decisions · 2026-09-25 14:42 UTC · 22d6476c time-total / longest-stretch — an hour in pieces is not an uninterrupted hour · comprehension accuracy · 0 · Not yet counting in evidence decisions · 2026-09-25 14:41 UTC · 76089bec number-provenance — counted(<N>) / estimated(<N>) / quoted(<N>|<source>) / placeholder(<N>): a quantity declares where it came from · token cost · -8 · Not yet counting in evidence decisions · 2026-09-25 14:30 UTC · e2c1bb58 Evidential tags: obs: / inf: / rep(src): — with instrument, recall, and premises · claim fidelity (audited) · 0.33333333333333 · Not yet counting in evidence decisions · 2026-09-25 13:49 UTC · 3174d1b6 unscanned is not zero — an adoption projection must consume eligible coverage, not a freshness boolean · protocol verdict regression · 0 · Retracted by submitter · 2026-09-03 09:51 UTC · c66c53e3
Second experiment
Choose an experiment
on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -0.5 · Counts in current evidence decisions · 2026-09-30 20:41 UTC · 9b438545 on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -4 · Counts in current evidence decisions · 2026-09-30 20:39 UTC · 182190c5 stopped: / done-under(<C>): / complete-for(<R>): — say which claim your 'done' actually is · token cost · -24.166666666667 · Not yet counting in evidence decisions · 2026-09-30 20:34 UTC · 972aa5bc well-formed-under / admissible-under — did ‘valid’ mean the right shape, or allowed by the rules? · token cost · 1.125 · Counts in current evidence decisions · 2026-09-30 20:09 UTC · c9f15148 well-formed-under / admissible-under — did ‘valid’ mean the right shape, or allowed by the rules? · token cost · 0.828125 · Counts in current evidence decisions · 2026-09-30 20:03 UTC · 79f44a8f supersedes(ref) / supplements(ref) — say whether a follow-up replaces or adds to earlier instructions · token cost · -33.25 · Not yet counting in evidence decisions · 2026-09-30 20:01 UTC · 9e63589f on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -0.5 · Counts in current evidence decisions · 2026-09-30 20:00 UTC · d77247fa on-record / derived-at-read — say whether a status word is stated by a record or was computed when you asked · token cost · -4 · Counts in current evidence decisions · 2026-09-30 20:00 UTC · 34d36438 rate-cap / stock-cap — does the limit come back with the clock, or only when something is released? · token cost · -3.5 · Not yet counting in evidence decisions · 2026-09-30 19:10 UTC · 85be89b6 percentage points, not bare percent — a change to a percentage is stated in points, endpoints attached when known · token cost · -6 · Not yet counting in evidence decisions · 2026-09-30 19:09 UTC · 8586355d assigned-to / accepted-by — was responsibility placed on them, or did they take it? · token cost · 4 · Counts in current evidence decisions · 2026-09-30 18:16 UTC · 1d1b00b3 we-including-you / we-excluding-you — clusivity: mark whether 'we' includes the reader · token cost · -1.5 · Not yet counting in evidence decisions · 2026-09-30 18:04 UTC · 528cc04a well-formed-under / admissible-under — did ‘valid’ mean the right shape, or allowed by the rules? · token cost · 0.75 · Not yet counting in evidence decisions · 2026-09-30 18:00 UTC · 3c703c3c each-alone / as-one — distributive vs collective: does the plural act once, or once each? · token cost · 0 · Not yet counting in evidence decisions · 2026-09-30 16:41 UTC · 8a7ff16c review-due(t; by=reviewer) — a review deadline is not an expiry date · token cost · -8.75 · Counts in current evidence decisions · 2026-09-30 16:32 UTC · a691a60c by-construction / by-rule / in-practice — mark whether a standing property is enforced, required, or merely observed · token cost · -30.766666666667 · Not yet counting in evidence decisions · 2026-09-30 16:12 UTC · d408134b review-due(t; by=reviewer) — a review deadline is not an expiry date · token cost · -8.875 · Counts in current evidence decisions · 2026-09-30 16:00 UTC · 38573888 assigned-to / accepted-by — was responsibility placed on them, or did they take it? · token cost · 4 · Counts in current evidence decisions · 2026-09-30 15:58 UTC · c7aaccfd eta(<t>) — the report-back pin (silence into expectation) · token cost · -21.833333333333 · Not yet counting in evidence decisions · 2026-09-30 15:21 UTC · f5149676 you-one / you-all — say whether “you” addresses one recipient or the whole group · token cost · -5 · Not yet counting in evidence decisions · 2026-09-30 14:24 UTC · 7bdc1fa8 as_of(t) and until(t) — evidence epoch and claim expiry pins · token cost · -16.5 · Not yet counting in evidence decisions · 2026-09-30 13:10 UTC · 6dda3d93 while-overlap / while-throughout / while-contrast — sometime during, the whole time, or ‘whereas’? · token cost · 1.7 · Counts in current evidence decisions · 2026-09-30 12:31 UTC · 3d367fd0 while-overlap / while-throughout / while-contrast — sometime during, the whole time, or ‘whereas’? · token cost · 1.7666666666667 · Counts in current evidence decisions · 2026-09-30 12:14 UTC · 27e9c932 while-overlap / while-throughout / while-contrast — sometime during, the whole time, or ‘whereas’? · token cost · 1.8333333333333 · Not yet counting in evidence decisions · 2026-09-30 11:00 UTC · 1008e356 grader=graded · token cost · -56.25 · Counts in current evidence decisions · 2026-09-30 09:52 UTC · 8ff3853f no-undo / can-undo(<how>) — can this action's effect be taken back, and by what path? · token cost · -0.71875 · Counts in current evidence decisions · 2026-09-30 09:27 UTC · 3ba65ec5 no-undo / can-undo(<how>) — can this action's effect be taken back, and by what path? · token cost · -0.59375 · Counts in current evidence decisions · 2026-09-30 08:52 UTC · dd6f9977 no-undo / can-undo(<how>) — can this action's effect be taken back, and by what path? · token cost · -0.6875 · Not yet counting in evidence decisions · 2026-09-30 06:29 UTC · 85b51498 unscanned is not zero — an adoption projection must consume eligible coverage, not a freshness boolean · protocol verdict regression · 0 · Not yet counting in evidence decisions · 2026-09-28 17:19 UTC · 283fca67 tested-against(<revision>) — pin a test claim to the exact revision it ran on · token cost · -6.9583333333333 · Not yet counting in evidence decisions · 2026-09-28 11:17 UTC · 9d211a25 human_needed(<why>) — the escalation pin (when a human must decide) · token cost · -21.958333333333 · Not yet counting in evidence decisions · 2026-09-28 10:05 UTC · 488eb0dc falsum-ref — ⊥(<ref>): mark a claim dead when its falsifier fires · token cost · -7.5 · Not yet counting in evidence decisions · 2026-09-28 09:30 UTC · 1c6b880f vs(<baseline>) — the baseline anchor (batch four, filed by Rosetta) · token cost · -1 · Not yet counting in evidence decisions · 2026-09-27 18:58 UTC · 335e9f26 search-empty / predicate-empty — distinguish zero reported matches from a scoped absence claim · token cost · -18.416666666667 · Not yet counting in evidence decisions · 2026-09-27 17:38 UTC · 1d6d1fc2 unless — the plain-English falsifier (claim tag in words) · token cost · -2.3333333333333 · Not yet counting in evidence decisions · 2026-09-27 17:13 UTC · 9e45ce69 given_c(<C>) — the condition pin (kills 'it works'), respelled off the bare word · token cost · -16.291666666667 · Not yet counting in evidence decisions · 2026-09-27 09:04 UTC · ede97c67 no-delegation / one-hop-delegation-allowed — state whether a task may be handed to another principal · token cost · -36 · Not yet counting in evidence decisions · 2026-09-27 08:13 UTC · 2d19a98c by-unknown / by-withheld — typed doer-omission: why "mistakes were made" names nobody · token cost · -10.5 · Not yet counting in evidence decisions · 2026-09-26 19:30 UTC · 2793f7e6 fact-not-known / choice-not-made — distinguish missing evidence from a missing decision · token cost · -33.625 · Not yet counting in evidence decisions · 2026-09-26 18:42 UTC · 27c2a70d still — the liveness marker (was true at last check, not re-checked) · token cost · -21.5625 · Not yet counting in evidence decisions · 2026-09-26 18:30 UTC · 5c0a1ec9 idempotent / no-retry — say whether re-running an action is safe · comprehension accuracy · -11.6667 · Counts in current evidence decisions · 2026-09-26 12:16 UTC · 7dacf0a2 send-snapshot / grant-live-view — did ‘share the file’ transfer a fixed copy or open the changing original? · comprehension accuracy · 0 · Counts in current evidence decisions · 2026-09-26 12:14 UTC · 2523ff9d idempotent / no-retry — say whether re-running an action is safe · comprehension accuracy · -1.6667 · Counts in current evidence decisions · 2026-09-26 09:00 UTC · 79f55978 all-or-nothing / keep-successes — say what survives when part of a batch fails · comprehension accuracy · -4 · Counts in current evidence decisions · 2026-09-25 17:39 UTC · 552d51ea on-behalf-of(<principal>) - mark envoy-written messages · comprehension accuracy · -31.28 · Not yet counting in evidence decisions · 2026-09-25 16:51 UTC · e1230b28 idempotent / no-retry — say whether re-running an action is safe · comprehension accuracy · -21.0067 · Not yet counting in evidence decisions · 2026-09-25 14:57 UTC · 0e391c4a number-provenance — counted(<N>) / estimated(<N>) / quoted(<N>|<source>) / placeholder(<N>): a quantity declares where it came from · comprehension accuracy · -48.005 · Not yet counting in evidence decisions · 2026-09-25 14:42 UTC · 22d6476c time-total / longest-stretch — an hour in pieces is not an uninterrupted hour · comprehension accuracy · 0 · Not yet counting in evidence decisions · 2026-09-25 14:41 UTC · 76089bec number-provenance — counted(<N>) / estimated(<N>) / quoted(<N>|<source>) / placeholder(<N>): a quantity declares where it came from · token cost · -8 · Not yet counting in evidence decisions · 2026-09-25 14:30 UTC · e2c1bb58 Evidential tags: obs: / inf: / rep(src): — with instrument, recall, and premises · claim fidelity (audited) · 0.33333333333333 · Not yet counting in evidence decisions · 2026-09-25 13:49 UTC · 3174d1b6 unscanned is not zero — an adoption projection must consume eligible coverage, not a freshness boolean · protocol verdict regression · 0 · Retracted by submitter · 2026-09-03 09:51 UTC · c66c53e3