Filed by PM seat domain:devx#2 (session_01VF48aw8RPG6wzDnMgp6rtw) while re-deriving #16468's blocker after #16465 landed (PR #21998 → 9c3bec0f4d, record 6019846983). Its measurements come from #21998's dev report (6018891729), and the seat re-read them on origin/main before filing. ⛔ Filed bare: grading and routing are triage's. ⛔ Not a claim.
Filing gate: ① a defect, class (a). A named producer's output contradicts measurement.
Measured
One run. On origin/main, the dataset's provenance.runs is ['37262126122'], and carriedOver is []. Every weight is a single observation. The refresh workflow's own header says the regeneration step "accumulates runs, each under its own --run <id> group … and medians the per-run sums across runs". This dataset carries one.
@objectstack/spec (test + test:repo summed on both sides, per #16550):
| source |
seconds |
ratio to the dataset |
refresh #20388 (ea7ff394b6, run 36380128221) |
1391.38 |
— |
refresh #21826 (run 37262126122), the current weight |
1134.86 |
— |
executed window, scheduled run 37433381795 |
1651.04 |
1.45× |
executed window, scheduled run 37453598388 |
1573.20 |
1.39× |
executed window, push run 37467882762 (70 executed, 0 replayed) |
1617.61 |
1.43× |
The seat read 37453598388's Test Core timing table itself (aggregator job 112248922952): spec 1573.20 against pinned 1134.86, and objectql 671.60 against 521.33 (1.29×).
@objectstack/objectql: pinned at 521.33 s, it reads 1.29–1.38× on six of nine post-refresh runs (#21998's report).
Run-to-run spread on one package: @objectstack/cli read 1033.74 s and 1723.05 s an hour apart (#21758's acceptance reading 6012200987). A single run is therefore a sample, not a baseline.
Why it matters now
Reader
Dedupe
MCP search_issues, repo-scoped, sorted by update: 「shard timings dataset refresh rests on a single run, spec under-recorded, median across several runs」 → 13 hits. One is open, #21933 (a merge-queue runner fault, unrelated). The closed ones include #21758 (the cli instance of this shape, fixed by the #21826 refresh), #16473 (the median merge rule for sliced packages), #16550 (the test:repo fold), #18341 (the refresh's write-back crash) and #16464 (the scheduled refresh itself). None records a dataset resting on one run, or spec's under-weight.
Dedupe words: test-shard-timings one run provenance · spec 1134.86 1573 1651 · ratchet ceiling 25% headroom standing red
Generated by Claude Code
Filed by PM seat
domain:devx#2(session_01VF48aw8RPG6wzDnMgp6rtw) while re-deriving #16468's blocker after #16465 landed (PR #21998 →9c3bec0f4d, record6019846983). Its measurements come from #21998's dev report (6018891729), and the seat re-read them onorigin/mainbefore filing. ⛔ Filed bare: grading and routing are triage's. ⛔ Not a claim.Filing gate: ① a defect, class (a). A named producer's output contradicts measurement.
scripts/test-shard-timings.json, written byscripts/measure-test-shard-timings.mjsthrough.github/workflows/shard-timings-refresh.yml, and last refreshed by chore(ci): refresh the Test Core shard-timings dataset #21826 (f2aa0c9fad).9c3bec0f4d, judges against it.Measured
One run. On
origin/main, the dataset'sprovenance.runsis['37262126122'], andcarriedOveris[]. Every weight is a single observation. The refresh workflow's own header says the regeneration step "accumulates runs, each under its own--run <id>group … and medians the per-run sums across runs". This dataset carries one.@objectstack/spec(test+test:reposummed on both sides, per #16550):ea7ff394b6, run36380128221)37262126122), the current weight374333817953745359838837467882762(70 executed, 0 replayed)The seat read
37453598388's Test Core timing table itself (aggregator job112248922952): spec 1573.20 against pinned 1134.86, and objectql 671.60 against 521.33 (1.29×).@objectstack/objectql: pinned at 521.33 s, it reads 1.29–1.38× on six of nine post-refresh runs (#21998's report).Run-to-run spread on one package:
@objectstack/cliread 1033.74 s and 1723.05 s an hour apart (#21758's acceptance reading6012200987). A single run is therefore a sample, not a baseline.Why it matters now
5924551316again, on a different package.Reader
RUN_COUNT: unbound variable—— 数据集已 22 天未动,而 pull_request 演练结构上跑不到这条腿 #18341, finding(ci): the refreshed shard-timings dataset records @objectstack/cli at 733s while whole-CLI runs measure ~1660s (2.27x), so Test Core shard 2 runs ~30 min and #16465 drift red cannot be wired #21758) sat indomain:devx.Blocked-by:written on it in the same act).Dedupe
MCP
search_issues, repo-scoped, sorted by update: 「shard timings dataset refresh rests on a single run, spec under-recorded, median across several runs」 → 13 hits. One is open, #21933 (a merge-queue runner fault, unrelated). The closed ones include #21758 (the cli instance of this shape, fixed by the #21826 refresh), #16473 (the median merge rule for sliced packages), #16550 (thetest:repofold), #18341 (the refresh's write-back crash) and #16464 (the scheduled refresh itself). None records a dataset resting on one run, or spec's under-weight.Dedupe words:
test-shard-timings one run provenance·spec 1134.86 1573 1651·ratchet ceiling 25% headroom standing redGenerated by Claude Code