docs(release): APR-RELEASE-001 spec + PMAT-1098 receipt for session 1 — the 0.67.0 train - #3164
Conversation
…c, and the R-17 model-movement row for PMAT-1098 The spec was dropped by the operator on 2026-09-12 and is untracked until this commit; the routing row records the orchestrator model moving opus->fable-5-1 at phase 4 (phase-boundary.sh). Pmat-Ticket: PMAT-1098
…(the 0.67.0 train) + the two agy lane receipts Andon obligation: K=240 was crossed at k_measured=312 when the operator issued the spec run; WIP committed and this draft PR carries the receipt-so-far. Amended at phase 5. Pmat-Ticket: PMAT-1098
|
§13.11 rung 1 — quorum shadow verdict Shadow mode: this records a verdict and merges nothing. A refusal |
…s the bottleneck, gx10 is 2.3x faster and half idle One append-only JSON per (sha, host, job) from the Actions REST API: queue_wait_s (created->started), exec_s, total_s, host, conclusion. peak_rss_mb and free_disk_gb are null with an explicit "[U]" in unmeasured[]: the REST API does not expose them, and an absent measurement must never read as a zero. What the 182 records say (p50 / p95 exec_s, this repo, 2026-09-12): workspace-test intel 1790 / 4148 s gx10 790 / 790 s yoga 259 / 3167 s guard-cargo intel 1201 / 1820 s gx10 422 / 516 s yoga 830 / 1046 s guard-tree intel 604 / 879 s gx10 192 / 213 s yoga 379 / 379 s queue wait p95 intel 2995 s gx10 233 s yoga 1654 s Per-run required-check wall clock, grouped by which boxes the run touched: gx10 only n=4 mean 9.4 min gx10 + yoga n=4 mean 11.3 min gx10 + intel n=10 mean 42.6 min gx10 + intel + yoga n=11 mean 31.6 min Every run that touched intel cost 32-43 min; no run that avoided it cost more than 11.3. intel took 15 of the 22 workspace-test records because that job is pinned to X64 -- and it is the job whose p95 is 69 minutes. #3139 unpins it. §1 coupling, no longer [U]: p95 gate = 72.3 min, so max PRs per train = 72h / 72.3 min = 59.7. §8's "stop cutting trains below 10" does not fire. Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
Build-server packing + build-time kaizen — 0.67.0 train, 12:35ZPacking right now (org runners, all repos; the aprender queue is what they are draining)
The finding: the packing target is being met on the wrong box. intel is at 94% 182 ledger records written this train (
Per-run required-check wall clock, grouped by which boxes a run touched:
Every run that touched intel cost 32-43 min. No run that avoided intel cost more §1 coupling, no longer Acted on this train
Still owed, and why it is not being done in this session: the P2 infra PR that |
… ledger answered §1 Three things the receipt did not yet carry: 1. orch_model. The harness moved this session from claude-fable-5-1 to claude-opus-5 in the middle of phase 4. A model change mid-session is a recorded event, never a silent one: route.sh record-event wrote the row to docs/audits/impl-routing.jsonl and the frontmatter now names the live model. fable_binding drops to false with the reason, rather than asserting a binding that no longer holds. 2. The build ledger, 182 records, and what they say about which box to fill. 3. §1's max-PRs-per-train, measured at 59.7 from 29 runs, so it stops being [U] and §8's stop condition can be evaluated instead of guessed. Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
…real defects, 38 classifier gaps The gate #3121/#3122 added during this cycle had never been executed against a tag. Its first run says 23 fail + 21 timeout, and that number is wrong in the interesting direction: re-running all 44 with each target's own required-features and keeping the logs leaves SIX genuine defects. 9 rows pass once the feature they name at runtime is enabled (cuda, compression, embeddings, tensor, shell-autocomplete). They declare no required-features and gate themselves in code, so the sweep ran them wrong. 20 rows are servers, TUIs and unbounded benchmarks. Killing them at 120 s and calling it `timeout` -- a class the skill defines as a defect -- is wrong on all twenty. 6 rows are a missing model or a missing argument worded outside the regexes. 1 is a TTY, 1 is this box's glibc against a prebuilt ort artifact, 1 is a wall-clock perf assertion failing under load. Verdict GO. Every one of the six defects is byte-identical at v0.66.0: the expect() on weight.rs:356, the five PinnedBuffer imports, the three cwd-relative config literals in llama2/train.rs. §3.1 skips a train whose CUT cannot go green; this cut did not cause any of them, the previous tag shipped them all, and a new instrument finding a backlog is not the same event as a regression. They are #3178 #3179 #3180 #3181, and the classifier gaps are #3182. Not claimed: that the 20 timeouts are healthy. They were not individually verified and the gate cannot yet tell a serving server from a wedged one. Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
… 34693750243..34700216664 incl. the discarded #3145 workspace-test (4958 s) and the two cancelled docs-PR runs [skip ci] Pushes to this docs branch were starting full CI runs on intel next to the train (34694025825, 34700216664, 34701097504 — all cancelled by hand); until v0.67.0 is tagged every push here carries [skip ci]. The final push before merge does not. Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
…e checkout paths — the shipped-tier machine-specific-path ratchet went RED on #3164 (+2 vs 81da9bc) [skip ci] check_hardcoded_paths.sh --full-if-capable, run 34700216664 step 43: evidence/release/0.67.0/agy/ph4-teamwork/delegate-receipt.json|/home/noah/src/aprender and |/home/noah/src/aprender/.git/config. Replaced with <repo_root>; the receipt is this session's artifact, not a tool output anyone re-reads by path. Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
… GPU hosts, cuda-nightly standing-RED since the v0.66.0 sha, why gx10 idles (X64 pin, #3139 795 s, ENOSPC), §6 triage to untriaged=0, one-at-a-time drain [skip ci] Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
…nder PR work under any intel queue pressure is P0 (selector row 0, §1, §3.4, §3.5 SSH, §5 P0·Pack + P0·Reap, §7 pack: line, §8 stops) [skip ci] Operator 2026-09-12, verbatim in §1. §3.4 no longer says one PR at a time: the queue builds 3 by design, dequeue only a KNOWN-RED group. §3.5 distinguishes durable config (forjar) from measurement/unclogging over SSH. Ground truth (§2) gains the measured X64 pin and the gx10/yoga disk layout. Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
…changed (SSH measurement, 146 GB reclaimed, forjar P0·Reap live on gx10+yoga via infra#549, packing rule, spec e874769), the stacked-merge-group resolver defect #3186 → #3187, 16 ledger records incl. fleet-pack samples [skip ci] Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
… on #3186, the GPU-host nightlies on the cut, fleet-pack samples [skip ci] Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
…or ruling 2026-09-13; the train publishes itself after dogfood GO + assets + preflight [skip ci] Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
…08 open / 424:124 arrival:closure / 95 fixed-but-open / 847 PR-less branches; five predicates, receipt in the ledger, no DONE without it (operator 2026-09-13, kaizen) [skip ci] Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
… (776 branches archived) / R4 1 / closure:arrival 22:61; 2 issues closed, 3 dirty PRs resolved, 1 closed, 10 verdicts pending [skip ci] Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
… — 3 bashrs NO-GO(s), then GO + tag at e45eaab Pmat-Ticket: PMAT-1098 Co-Authored-By: Claude Fable 5.1 <noreply@anthropic.com>
APR-RELEASE-001 §7 — session reportTag: https://github.com/paiml/aprender/releases/tag/v0.67.0 · release commit 🤖 Generated with Claude Code |
T-5 R-4 — verdicts applied (2026-09-13T10:40Z)
Also: mini-m4 re-registered and taking The reconcile receipt file update rides the next docs PR (this one is in the merge queue; a push would dequeue it). |
Draft until the train reaches its terminal step. Carries the operator's spec (untracked until now), the PMAT-1098 receipt-so-far (andon: K crossed at k=312), the R-17 routing row, and the two agy lane receipts under evidence/release/0.67.0/agy/. The §7 report and the docs/build-ledger train record are added when v0.67.0 is tagged or SKIPPED.
Refs #3078, PMAT-1098.
🤖 Generated with Claude Code