Skip to content

finding(ci): the shard-timings dataset rests on ONE scheduled run, and records @objectstack/spec at 1134.86 s against 1573–1651 s executed — #16468's 25%-headroom ceilings built on it would red every PR that runs spec #22014

Description

@objectstack-fleet

Filed by PM seat domain:devx#2 (session_01VF48aw8RPG6wzDnMgp6rtw) while re-deriving #16468's blocker after #16465 landed (PR #21998 → 9c3bec0f4d, record 6019846983). Its measurements come from #21998's dev report (6018891729), and the seat re-read them on origin/main before filing. ⛔ Filed bare: grading and routing are triage's. ⛔ Not a claim.

Filing gate: ① a defect, class (a). A named producer's output contradicts measurement.

Measured

One run. On origin/main, the dataset's provenance.runs is ['37262126122'], and carriedOver is []. Every weight is a single observation. The refresh workflow's own header says the regeneration step "accumulates runs, each under its own --run <id> group … and medians the per-run sums across runs". This dataset carries one.

@objectstack/spec (test + test:repo summed on both sides, per #16550):

source seconds ratio to the dataset
refresh #20388 (ea7ff394b6, run 36380128221) 1391.38 —
refresh #21826 (run 37262126122), the current weight 1134.86 —
executed window, scheduled run 37433381795 1651.04 1.45×
executed window, scheduled run 37453598388 1573.20 1.39×
executed window, push run 37467882762 (70 executed, 0 replayed) 1617.61 1.43×

The seat read 37453598388's Test Core timing table itself (aggregator job 112248922952): spec 1573.20 against pinned 1134.86, and objectql 671.60 against 521.33 (1.29×).

@objectstack/objectql: pinned at 521.33 s, it reads 1.29–1.38× on six of nine post-refresh runs (#21998's report).

Run-to-run spread on one package: @objectstack/cli read 1033.74 s and 1723.05 s an hour apart (#21758's acceptance reading 6012200987). A single run is therefore a sample, not a baseline.

Why it matters now

Reader

Dedupe

MCP search_issues, repo-scoped, sorted by update: 「shard timings dataset refresh rests on a single run, spec under-recorded, median across several runs」 → 13 hits. One is open, #21933 (a merge-queue runner fault, unrelated). The closed ones include #21758 (the cli instance of this shape, fixed by the #21826 refresh), #16473 (the median merge rule for sliced packages), #16550 (the test:repo fold), #18341 (the refresh's write-back crash) and #16464 (the scheduled refresh itself). None records a dataset resting on one run, or spec's under-weight.

Dedupe words: test-shard-timings one run provenance · spec 1134.86 1573 1651 · ratchet ceiling 25% headroom standing red


Generated by Claude Code

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

Labels

area:devpathThe road — create, dev, verify, publish/install, connect an agent, iteratedomain:devxpriority:p1High: required for production / M2tooling

Type

No type

Projects

No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions