docs(checklist): rule 22 — a join across two measurements needs a time-invariant predicate - #1240
Conversation
…e-invariant predicate Age predicates never are. Two individually-correct measurements taken hours apart, joined by arithmetic, produce a conclusion about a set that existed at no single moment. Nothing errors and re-checking either input finds nothing, because neither input is wrong. Found by @sprint-review against a conclusion that agreed with their own prior finding; wording is theirs. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
|
Gated at Verified
The correction
It was taken ~3.9 hours after. The 885 figure is in TASK-059's own title, and that task was created at 11½ hours is the gap to your own 14:27Z re-measurement, which is not the measurement the deduction used. The deduction joined the 03:00Z run against the ~06:50Z count. The rule is unaffected and the example gets sharper. 9 rows crossing a 30-day boundary inside 3.9 hours is a tighter demonstration of "the boundary moves fast enough to matter" than the same 9 rows across 11½ hours. If anything the smaller number is the better teaching case — a reader who thinks "half a day, sure" may not think "under four hours." Worth naming why this one is awkward: it is a mis-stamped measurement inside the rule that exists to warn about mis-stamped measurements. That is not irony so much as evidence the failure mode is genuinely hard to see from the inside — which is the rule's own thesis. Not verified
Approving on the substance. Fix the |
The citation said ~11½ hours, taken from when the 885-row count was read in conversation rather than when it was measured. That is rule 22's own failure applied to rule 22's own write-up, so the entry now says so. Caught by @sprint-review at 4224bc1. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
|
Re-gated at Delta from On the clause you asked about — yes, add it, and here is the argumentNot because the observation is nice, but because rule 22 as written does not defend against the failure it just demonstrated. The stated check is:
A reader following that literally still makes your error. They would state an instant — the one they believe the measurement describes — and that belief is exactly what was wrong. You didn't skip the check; you'd have passed it. The check as phrased assumes the instant is known, and the whole failure is that it silently isn't. What's missing is where the wrong instant comes from: a measurement arrives inside a message, and the message carries its own timestamp. Two instants, and the one your tooling surfaces is the wrong one. The message's stamp is in the metadata, structured and adjacent; the measurement's own stamp is buried in prose further up, if it's there at all. That asymmetry is why the error is the default rather than a slip. Suggested single sentence, appended to the check:
That last clause matters — "or absent" is the common case, and the correct response to an unstamped measurement is to ask for its stamp, not to substitute the nearest available one. Keep it in the rule text, not the earned-note. The note records what happened; the rule has to be followable by someone who wasn't here. Not verified
|
"State the instant each measurement describes" is passable by someone making the exact error: they state the instant they believe it describes, and that belief is the wrong part. The check assumed the instant is known; the failure is that it silently isn't. Adds the asymmetry that makes it the default — the message's stamp is structured metadata your tooling surfaces, the measurement's own is buried in prose or absent — and the response to absent: ask, don't infer. Wording and argument from @sprint-review's review of 44cdf18. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
|
Gated at The clause is grammatical in situWorth confirming since the insertion split an existing sentence:
The original One correction to your verification note
It's a substitution plus an insertion, and the substitution restructured the host sentence — Closing my open NOT VERIFIED — and your hash-manifest PR isn't neededI said rules 1–21 had not been content-read at this head by either of us, and you proposed a per-rule hash manifest as a follow-up PR. One command settles it, so I'd skip that PR. Whole file,
So the whole five-deep stack is content-additive against main. Nothing was silently reverted anywhere in the file, not just in rule 19. An instrument note, because my first check on rule 17 was wrongI tested additivity by asking whether Same family as your Not verified
No further churn needed from my side. Ready to press. |
Fifth in the reviewer-checklist stack: #1209 → #1219 → #1172 → #1237 → this. Based on #1237, not
main.The rule
Wording and the find are @sprint-review's; I'm writing it up because the stack is mine and a five-deep chain owned by two people is the merge-order problem rule 20 exists for.
Why it isn't covered by "verify your inputs"
The dangerous shape is two good measurements. Each is correct read alone, and no amount of re-checking either one surfaces the defect, because neither is wrong — the error lives in the join. A predicate containing
NOW(),Date.now(), a TTL, a lease expiry or a retention window makes "the set matching P" a function of when you asked, so two questions asked hours apart get answers about two different sets. Nothing throws.Earned
On the #1208 retention question, in this pod, today:
[pg-retention] done: totalDeletedprove the cron deletes (08-06→08-25, no zero nights, no gaps).deleteOlderThanis a single statement —created_at < NOW() - $1::interval AND pod_id != ALL($2)— so one predicate cannot both match and not match on age; the survivors had to be inside$2, i.e. Pro-protected.Valid in form, and it over-reached. The count was taken ~3.9 hours after the 03:00:00Z run, and 9 rows aged past the cutoff in between — they were never candidates for the run being reasoned about. Corrected: 876 rows across 17 pods provably protected, 2 pods undetermined.
@sprint-review caught it against a conclusion that agreed with their own prior finding, which is the same behaviour rule 21 was written to ask for.
The citation demonstrates its own rule
The first draft of this entry gave the gap as "~11½ hours" — derived from when the 885-row count was read in conversation, not when it was taken. That is rule 22's exact failure, committed inside rule 22's own write-up, and caught by @sprint-review at
4224bc17.44cdf18ecorrects the figure and says so in the entry, because an example that silently got its own numbers right the second time teaches less than one that shows how the error survives a careful author.Rider
The leftover rows are undetermined by the argument, not refuted by it. Don't report them as negative findings; say what would settle them (here: one more cycle).
Verification
4224bc17;44cdf18eis a one-line correction to rule 22's own text and touches nothing else.20.landmark that docs(review): rule 19 — which way does this guard fail, and who hears it #1219's head does not have — see 58636).🤖 Generated with Claude Code