- wave D (48 adapt cards): 47/48 standalone, 45 BANKED across two gates (40 + 5 late-repair). Operational lesson: build the gate slate AFTER the repair stage lands — the first slate was built early and 5 rescued drafts needed a second gate. - wave C (35: 11 tell + 24 weak): 32 banked, 91% gate. Weak lane proven 24/24 on haiku. - MAIN BLOCKER DIAGNOSED (the night's most valuable finding): main drafts are byte-correct yet gate 0/4. Byte-diff of the built EXE = exactly 2 bytes in 413,696, NOT in the drafted fn: a jal at 0x80060E74 retargets func_80061FA8 (game code, 800c2.o) -> firstfile (PsyQ libapi A66.o). Adding one C fn perturbs symbol resolution between game code and the LINKED PsyQ archives. main is an INTEGRATION wall, not a matching wall -> its own lane; excluded from build_wave_atlas by default. - build_wave_atlas.py now packs by (binary, TU) — the REAL gate-group key, since each group is one whole-binary rebuild (wave D: 42 drafts / 23 groups = the throughput ceiling). - 3 NEARs enqueued as grinder fuel incl. func_80183578 at close=1 DELAY-SLOT (§60a precedent).
34 KiB
CURRENT PHASE — Phase 31: The Frontier Atlas & Wide-Tolerance Campaign
Started: 2026-08-14 · Plan approved: 2026-08-14 (gate 1; Drew) · Effort doctrine: xHigh default / Max deep (T5, T7, synthesis) / Ultracode waves (R26/R27 prompts) / Fable-tier only for new wall classes.
Approved plan: /home/musashi/.claude/plans/fable-5-set-max-goofy-seahorse.md (the full design; this file is the crash-recovery log).
Approval also ratified R37 (probe before costing), R38 (read recorded failure verdicts first), R39 (negative-control new refusal-checks) — now binding.
The re-charter (one paragraph)
Instead of roadmap-v2 P31's per-function grind, Phase 31 organizes the 12,059 remaining stubs (main 1,041 · resident 14 · ov 9,832 · md 1,179) into crack groups: a deterministic per-function feature layer + multi-tier similarity atlas (li-normalized exact tier, seed sweep vs the 2,719 matched skeletons, calibrated warm tier for the 3,238-unit cold tail + main, kNN neighborhood graph), evidence-joined to a lever label per group; plus widened mechanical lanes (LEN-tolerant aligned remap with a fully-mechanical LI class, §172b EXTPAIR/SELECT detectors + routing, PLUMBING campaign, permuter cluster warm-start, weak-seed cards); then a campaign loop to ceiling — deterministic $0 lanes first, agents only for exemplars, velocity-ledgered, closed on measured decay. Milestone shape = P30 (campaign to ceiling; every remaining stub on a named ledger at close). Main fully included from day one.
Task checklist
- T0 — Pivot log + freshness + hygiene — DONE 2026-08-14. (xHigh)
- T1 — Integration quick-bank sweep — DONE 2026-08-14 (pending final R22 log line). 8 banked, 0 agent tokens. (xHigh)
- T2 — References — DONE 2026-08-14. (xHigh)
- T3 — Main enablement — DONE 2026-08-14. (xHigh)
- T4 — atlas_features.py — DONE 2026-08-14. (xHigh)
- T5 — atlas.py — DONE 2026-08-14. THE ATLAS EXISTS. (Max)
- T6 — PLUMBING campaign — DONE 2026-08-14. +9 banked (probe 64%); recipe + 3 laws distilled (§173); sweep proved the no-draft majority routes to family lanes. (xHigh)
- T7 — family_align.py built + NC'd; the mechanical-cousin premise REFUTED by its probe (0/26) — engine re-scoped to T8's LEN+N same-function pile; no driver built (correctly). (Max)
- T8 — LEN+N lane — DONE 2026-08-14. Pile routed 587/587; 345 wrong-drafts reclassified; 49 permuter + 192 card fuel staged; mechanical lane = honest null. (xHigh)
- T9 — Warmstart + weak-cards — DONE 2026-08-14. Grinder queue armed (49 lenmiss + 10 seeded drafts, 120 refused by the stream filter); 954 weak-cards emitted, cheap-tier-dominated. No grinder patch needed. (xHigh)
- T10+ — Campaign loop to ceiling (repeating sessions; velocity ledger; close on measured decay). (Ultracode waves / xHigh orchestration / Fable new-walls)
- Tclose — PhaseEnd (gate 2). (Max)
Standing verification (every task)
R22 clean-fleet 213/213 after every banked batch · tools-health green · 0 NON_MATCHING (G4) · dedup-check 0 failed · R32 coverage assertions on every new scanner · R39 negative controls on every refusal check · R37 probes before pricing · R38 ledgers before experiments · commit per task (task + this log in the same commit; Drew pushes).
Progress log
-
2026-08-14 — Phase planned and approved (3 Explore + 2 Plan agents; full design in the plan file). Task list built (harness tasks #1–#12). T0 started.
-
2026-08-14 — T0 COMPLETE. (1) R31 decision-log entry (the re-charter WHY + R37–R39 ratification). (2)
harvest_verify.pyimport guard: a bare import now RAISES instead of running a gate (verified both directions; CLI behavior unchanged). (3) Resident ±1 RESOLVED + FIXED:--bootstrap's linear partition had fused the +0 data word withfunc_800CEDFC(row0x800CEDF8nins=18) and droppedfunc_800D33E0past a glued tail — the true denominator is 145 (progress was right, the sig wrong).sig-residentnow ELF-seeds (S45 pattern: unique 4-aligned T-symbol addrs inside theresident_TEXT_START/ENDmarkers → exactly 145; bootstrap fresh-clone fallback). All three oracles now agree (sig 145 · corpus matched 131 · progress byte-ident 131);audit-corpus0 PHANTOM + 0 TRUNCATED. (4) Family maps regenerated at HEADcommit:2161: 11,025 open non-main members = 12,059 − main's 1,034 EXACT (the stale map's 102 phantoms cleared); cousins totals now A-prop 1,125 / seeded 1,293 / cousin-multi 5,373 / cold 3,234 inst; adapt cards 704, aprop cards 204 (full emission). (5) Main fuel gap is DEAD: 2,001/2,002 main stubs have cached Ghidra-C (onlyfunc_80049600missing) — the roadmap's "0/2,096" note was stale. (6)make tools-health→ OK (dedup 2,063/0; C1 254,521/254,521; audit-digest confirms the fleet digest; resident fix moved instr num+denom by the same +9). -
2026-08-14 — T1 COMPLETE: 8 banked for 0 agent tokens. R38-first: partitioned the MATCH-108 pile against current stubs (75 still open) and against the S50 gate history (62 gated-and-failed with verdicts · 13 never-gated). The lanes and their measured outcomes:
- Never-gated 13 → gate_lane: 0/13, but the verdicts decomposed to 11×
undefined reference to D_*= the §171 stale-seed-symbol class. Extendedaprop_symfixwith STALE-DELTA (n:n uniform-delta rebase; R39 synthetic + snapshot NCs, zero false positives; the delta test even refused a pair my hand-check wrongly accepted) → 4 rebased, 4/4 banked (func_8016BCC0,func_8017F1C8,func_80186BD8,func_80186BF8). Cookbook §171-D written in-session. - SELF-decl PLUMBING 7 →
recover_integration --stages demacroize --max-tier binary: 4/7 banked (func_80139BE0,func_8014ED28,func_80161D88,func_801659DC); 3 stay near. - Stored-draft re-gates (no-verdict 7 + close=0 8 + diff 1 + 9 STALE→clean world-motion drafts): 0/23-ish banked — the ~8% A10 stored-verdict law held again; all re-verdicted fresh.
- Handed forward with fresh classifications: CALLEE-decl 15 + CC1-FAIL 14 → T6 (cast-callees/tu-scope stages); UNDEF-DATA/OTHER 9 → the §171b-1 data-definition carry (T7/T8); md CARVE-REFUSED 8 → campaign side-quest ledger. immfix pile: fully consumed (0 open). fix20: 19/20 consumed in S50.
- Tool fixes landed:
gate_lanepropagate-commit tag now derives from GATE_PHASE (was hardcoded phase-30 S49). - Rate lesson for the velocity ledger: fresh-fix lanes (STALE-DELTA 4/4, demacroize 4/7) vastly outperform blind stored re-gates (0/23) — the campaign loop's L2 ordering is confirmed by measurement.
- Never-gated 13 → gate_lane: 0/13, but the verdicts decomposed to 11×
-
2026-08-14 — T2 COMPLETE (references). (1) PsyQ dev-CD extracted: walked the on-disk Track-1 image (MODE2/2352) with the frozen
tools/bfm_extract/iso9660.py(R33 — no new extractor; walker = iter_directory/read_extent with out-of-range extents skipped) →tools/reference/psyq-sdk/(gitignored): 2,374 files / 231.6 MB, 400 C sources (373 in PSX/SAMPLE/ across CD/GRAPHICS/SOUND/MODULE/CMPLR/…); only 7 out-of-track.DAaudio skipped. (2) Provenance find:GNU/SNGNUVER.TXT= SN Systems' gcc build history (2.7.2.SN32.3.7.0002, 14.5.97) with per-build changelog of SN's patches vs vanilla — onlyUNROLL.C(parameterised max unroll insns) is codegen-relevant; recorded in the idiom notes as the first-look suspect if a loop-unroll residual ever defies the vanilla model. (3) gcc-2.7.2 reference completed: +6 files from GNU ftp (calls.c+caller-save.c— both cited by §172's producer model, previously missing — + integrate/optabs/varasm/recog), tarball sha2567cd8bce5…recorded. (4)docs/psyq-sample-idioms.mdseeded (inventory, provenance, first style conventions, the lane hook); SETUP §5.6 rows added (R21). -
2026-08-14 — T3 COMPLETE (main enablement). (1)
sig_imageseed-ends extension:--seedsnow accepts0xADDR NINSlines (and jsonlnins); a seeded nins is authoritative — the slice is exactly[addr, addr+4·nins), bypassingfunc_end(whose heuristic mis-sliced 3/40 main samples). (2)corpus.s_ins_count()factored from audit() (R33, one counter) +corpus.py <bin> --seed-endsCLI. (3)make sig-main: 2,002 main stubs signed at splat-true lengths →.run/sig.main.jsonl(deliberately splat-SEEDED — the atlas needs the boundaries a match must hit; NOT the independent second oracle, which stays deferred perdocs/second-oracle.md). (4) Full word cross-check: 2,002/2,002 slices byte-faithful (EXE bytes vs.scomment-column words; 0 SLICE-SUSPECT; note the.sword field is byte-order hex, not the LE-decoded value — first checker draft compared wrong and read 1,999 false suspects). (5)family_remapmain special-cases:vram_of('main')= code-segmentvram − startfromsplat.us.exe.yaml= 0x8000F800;img_path('main')from the same yaml's target_path (R33).stream_words('main')==.swords on 25/25 samples. (6) Regression:sig-residentre-run byte-identical after the sharedread_seedschange. -
2026-08-14 — T4 COMPLETE (
tools/atlas_features.py). Per-function features memoized per distincth_exact(.run/feat_memo.json) and fanned out 1:1 with sigs to.run/feat.<bin>.jsonl×213 (main reads the splat-truesig.main.jsonl; registry untouched). 363,525 rows / 92,855 distinct bodies in 21 s (vs the 3–6 min estimate). Features: nins/band · o0 prologue tell · frame/callee-saved set+order/fp · CFG skeleton (nblk/ncond/nback/ret_n via branch-target scan; jal = call, never an edge — position-independence proof in the docstring) · mid_jr/jalr · stable-call-sequence hash (fixed main+resident ranges) · reloc-kind sequence hash · 16-bucket ophist · §172b tells as importable shared functions (extpair/dupselect/sign_mix/magic_div — T8's len_tells imports these, R33) · li-normalized skeletonh_seqn(non-anchor lui dropped, ori→addiu class). Verifies: R32 join asserted per binary at write; determinism 0 mismatches on 200 re-derived bodies; mid_jr cross-check vs family_hseq's independent oracle: 6,444/6,444 exemplars agree (first verifier draft compared ZERO rows — a string-vs-int addr type mismatch, the R32 silent-no-op class caught in my own verifier; now indexed + fails loud if compared==0). o0-tell vscorpus.is_o0on 13,019 open fns: 41 both / 94 tell-only = candidate undiscovered -O0 fns in -O2 TUs (§116 — o0_subsplit carve fuel for the campaign) / 14 src-only outliers. -
2026-08-14 — T5 COMPLETE (
tools/atlas.py+make atlas) — THE FRONTIER ATLAS EXISTS. Survey in 92 s, all assertions green: 5,139 groups cover ALL 12,058 open instances / 613,710 ins (12,058 = progress's 12,051 stubs + 7 NON_MATCHING exactly — the +1-blob and +8-banked discrepancies were both chased and resolved: data blobs excluded, T1 banks confirmed absent; main open reconciles to classify()'s stubs+NM with a hard assert). Tiers: calibration measured (positives = cousin ≥0.85-raw merges, negatives = size-matched cross-unit pairs; the li-normalized metric is so discriminative that recall ~99% holds down to 0.55; rule = smallest t with neg-accept ≤0.2% ∧ recall ≥95% → THRESH_WARM=0.70 at 99.1%/0.18% — chosen because a false warm-merge wastes an exemplar crack, a miss only routes cheaper) · T1.5 h_seqn 2 merges · warm tier 1,019 merges · seed sweep: 4,702/7,247 open skeletons (65%) carry a ≥0.55 matched seed from the 2.7k-skeleton pool · kNN graph (top-8 open + matched neighbors) · tiny-direct. Lever table (the strategic map): head-crack 1,283g/186.9k ins · UNKNOWN 1,964g/138.6k (honest) · extend-tell 575g/76.7k · redraft 46.8k · jtbl-carve 190g/45.7k · integration 575 inst/23.4k · seeded-crack 23.4k · len-vein 778 inst/16.8k · swaprepeat 9.2k · plumbing 161 inst/8.1k · o0-lane 6.6k · cc1 6.4k. Top group: 268 drifted single-member skeletons unified (per-location ~23-ins family, 0.72 seed). Evidence joins accounted (audit/backlog/ledgers/cards; residuals EXCLUDED-STALE until regenerated);--targetsresolves 12/12.s.make atlas= the full regen chain. docs/frontier-atlas.md committed. -
2026-08-14 — T6 IN PROGRESS (PLUMBING campaign) — probe converged after two R14 corrections.
plumbing_groups.pyderived the honest pool: 237 still-open PLUMBING rows (the "1,217" was ledger-vintage inflation): SELF 109 (3 concentrated binaries) · CALLEE 48 · OTHER 48 · DATA 32. Probe (ov_SC03_107 SELF-21 → 14 with drafts): first run 0/14 with a PHANTOM shared error — root-caused to cross-group TU-edit poisoning (a demacroize edit in TU-A persists while TU-B's drafts gate; every whole-binary build compiles ALL TUs) → fixedrecover_integrationwith per-group isolation (git-checkout binary TUs between groups, gate_lane's proven pattern; engine_core.h untouched so fleet-tier arity persists) + added themacro-externsdraft stage (§121: one draft's guessedextern int f()vs the TU'sDEFINE_-macro definition — lifted from family_sweep, R33) +tu-scopestage (§103 STU, the sweep-only lever). Isolated re-run exposed the TRUE class:undefined reference to D_*= the §171 stale-seed-symbol class (rtu_match MATCHes these drafts — it is blind to relocation names; the R34 two-oracle disagreement exactly as documented) → symfix-first: 11/11 rebased (mixed 1:1 + n:n deltas incl. the S49 Δ0x4128) → 9/14 banked (64%) where the raw path scored 0. The standing recipe:aprop_symfix --fix → recover_integration --stages macro-externs,demacroize,tu-scope(per-group isolated). Full sweep over the remaining ~200 rows now running. -
2026-08-14 — T6 COMPLETE. Sweep over the remaining ~200 rows: the decisive finding is Law 3 (§173) — the biggest groups (ov_SC02_037 44 rows, most of ov_MAIN_012's 42) have zero stored drafts: their PLUMBING verdicts came from transient family-sweep remaps never persisted to the backlog. A recovery lane can only recover what was stored — verdict-only rows are family-lane fuel (already atlas-labeled), not recovery fuel; the true stored-draft class was largely consumed by the probe (9 banked, 64%). Singleton tail with drafts: 0/~20 (fresh per-fn verdicts recorded). Cookbook §173 (symfix-first · per-group isolation · verdicts-without-drafts) + index regenerated (518). R22 clean fleet: 213/213 with all T6 banks. Phase total so far: 17 banked, 0 agent tokens (stubs 12,059 → 12,042).
-
2026-08-14 — T7 COMPLETE (honest refutation).
tools/family_align.pybuilt: aligned classifier (SequenceMatcher over FC.tok; LEN-LI/LEN-NOP/LEN-JTBL/LEN-STRUCT/STRUCT-ALIGNED/PURE/IMM verdicts) + li-cluster reconstructor (with the split-cluster absorb for the rs-changed addiu partner) + the aligned imm engine. NC-1: 157/157 verdict-equivalence with classify_member on banked pairs — the NC itself caught two real classifier gaps (R-type non-shift sa = STRUCT; registers tested BEFORE the reloc skip — a reloc-slot word with a different register is STRUCT). NC-2 parity 21/21. The R37 probe then killed the planned driver before it was built: 0/26 LI-ONLY cards classify mechanically (regfields drift ×19) — cousins are 0.85-similar DIFFERENT functions; §168 law 1 ("a cousin is a seeded crack, never a remap") re-derived by measurement. The engine's true consumer is T8's LEN+N pile (draft vs its own target = same function). Parked for T8: the reloc-vs-constant range discriminator for lui-bearing clusters. Decision-log entry written (R31). -
2026-08-14 — T8 COMPLETE (LEN+N lane). Built:
match_one --emit-streams(additive, stdout-identity NC'd) ·family_align.addr_true_rel(the reloc-vs-constant range discriminator — full conservative set kept for pair semantics/NC-1 [157/157 regression], address-true subset for indel eligibility; synthetic probes: lui-bearing constant cluster now LEN-LI, address-anchored indel still refuses) ·tools/len_tells.py(per-draft aligned classification + §172b tell tagging on target-side indels, detectors imported from atlas_features R33, cookbook text embedded in cards) ·tools/lenmiss_route.py(pool-parallel per A8: 587 audit LEN rows re-verified live and routed in 24 s). Routing: redraft 345 (frac>0.35 — the draft is not the function; reclassified APPEND-ONLY in the backlog, so near-miss metrics stop lying) · permuter-length 49 (|Δ|≤2 clean drift — grinder fuel) · cards 192 incl. 14 EXTPAIR/SELECT/NOP-tagged (the audit's own detectors had emitted zero — these are new signal) · mechanical 0 — an honest null: stored drafts rarely get CONSTANTS wrong (agents copy them from asm); LEN drift is shape, so the LI-cluster swap lane has no fuel in this pile and family_align's value here is the classifier/detector. R32 accounting 587/587. -
2026-08-14 — T9 COMPLETE.
tools/warmstart.py(the permuter/grinder FEEDER):--from-bankedwalks a banked exemplar's h_seq family's still-open members, builds remapped proven-body drafts (symbol_map+aprop_autodraft.build_draft, refusing on reloc-count mismatch), stream-classifies member-vs-seed with ZERO compiles (masked_diff-shaped dicts from ground-truth bytes →residual_class.classify_streams), enqueues ONLY permuter-shaped (bucket==permuter or LENGTH-DRIFT |Δ|≤2) as backlog near-records;--lenmissingests T8's 49-route. Grinder patch NOT needed (its candidates() deliberately keeps unclassified records — "unknown is not a reason to skip" — so pre-filtered enqueues flow as-is; documented in the feeder's docstring). Armed live: 49 + 10 enqueued, 120 refused by the stream filter (the anti-92%-wasted-CPU discipline working).family_cousins --weak-cards: 954 units (the 0.70–0.85 annotate-only band, never before consumed) as seeded-crack cards, ins-ranked, §168 laws embedded, model-routed haiku 804 / v3 43 / sonnet 86 / opus 21 (cheap tiers dominate — the token-efficiency shape), 0 unresolved.s.
Campaign velocity ledger (T10+)
| wave | lane | cards | standalone MATCH | banked | tokens | tok/bank | lesson banked |
|---|---|---|---|---|---|---|---|
| A | adapt SMALL-EDIT | 24 | 18 (75%) | 12 (+1 prop) | ~1.40M | ~117k | zero stale seed-symbols (the §171 prompt-law works); 6 gate-fails all INTEGRATION shapes (3× decl-type vs TU, 1 arity, 2 TU-context DIFF) → wave-B prompt adds match-the-TU's-existing-decl |
| grinder-1 | permuter | 10 | — | 0 | $0 | — | func_800CB4CC parked at best-1 (warmstart seed); queue enriched +59+6 records since |
| C-probe | tell (3) | 3 | 1 (33%) | 1 | ~300k | ~300k | §174 Law 4: the TU's decl of YOUR OWN fn constrains the def sig — standalone MATCH gated 0/1 until canonical-sig + cast-at-use; tell attribution unreliable (2/3 residuals were a different class) |
| C | tell 11 + weak 24 | 35 | 35 (100%) | 32 (8 propagated ×2) | ~2.39M | ~75k | Law 4 in-prompt → 0 symbol fails, 91% gate; weak lane 24/24 on haiku — the 890-card vein is live; 3 NEAR → grinder |
| E-probe | mass/main (6) | 6 | 4 (67%) | 0 | ~474k | — | main is agent-draftable but LINK-BLOCKED: 1 byte-correct fn ⇒ 2-byte whole-EXE diff, a jal retargeted game-code→PsyQ-archive symbol. Not a matching wall. gate_lane main-blindness fixed en route |
| D | adapt (48) | 48 | 47 (98%) | 45 (40 main gate + 5 late-repair, 2 gates) | ~4.48M | ~100k | repair stage rescued 5/6 first-pass DIFFs → gate them in a SECOND slate (the first slate is built before the repair stage lands); 23 gate groups for 42 drafts = the throughput ceiling → build_wave_atlas.py now packs by (binary,TU) = the real gate-group key; 1 NEAR at close=1 DELAY-SLOT → grinder |
-
2026-08-14 — CAMPAIGN OPEN (T10+, Ultracode). Wave A: 24 adapt cards → 18/24 standalone (75%) → symfix-first (0 stale — the prompt-law worked) → gate 12 banked + 1 propagation (~117k tok/bank e2e; stubs → 12,030). 6 NEARs (3 at close ≤3) enqueued to the grinder. Wave B (48 cards, all haiku, 29 binaries) launched with the decl-matching lesson. Grinder pass 1: 0/10 but func_800CB4CC at best-1.
-
2026-08-14 (resume session) — Checkpoint resume steps 1–2 DONE. (1) R22 owed proof banked:
make clean && make extract-all && make check-all→ 213 passed, 0 failed of 213 (.run/r22_p31_resume.logEXIT=0) — wave B's banks verified fleet-clean. (2)make atlasregenerated at HEADcommit:2234: 363,525 feature rows / 92,855 distinct bodies (0 newly computed — memo hit); 5,144 groups / 11,994 open instances / 612,325 ins; warm merges 1,014; seeded 4,710/7,247 skeletons; all assertions green. Open count 12,058→11,994 reconciles with the session's 65 banks. (3) Wave C fuel verified:.run/wave_p31c_cards.json= 38 cards (14 tell/sonnet + 24 weak/haiku); grinder idle (heartbeatdone); gate free. Drew's re-extraction question answered (disc/splat layers gain nothing — deterministic + continuously regenerated; sig/atlas layer regenerated by this step; Ghidra-C re-analysis = targeted-probe candidate only, R37). NOTE: this harness build has no TaskCreate tool — this checklist + log is the monitorable task view (R28 noted plainly). -
2026-08-14 (resume session, cont.) — WAVE C COMPLETE: 32 banked, R22 213/213. Probe→wave cadence (§174 law 3) paid: the 3-card probe cost ~300k tok and returned Law 4, which the 35-card wave then converted at 91% gate with zero symbol failures. Independent re-verify of all 35 drafts by me (R14) before gating: 35/35 MATCH held. Reach measured honestly: 32 exemplars, only 8 had any sharer, each ×2 → ~1.25× effective. The ×134 era is over (P25/29/30 harvested the shared cores); fleet-% now moves ~1:1 with exemplars banked, so THROUGHPUT is the lever, not leverage.
- THREE instrument defects found+fixed in my own new tooling this cycle (the R32/R35/R39 class, and the reason R39 exists): (1)
build_wave.py's gate-guard usedpgrep -fviashell=True— the wrappingsh -ccarries the pattern in its own cmdline so it self-matched and refused forever; fixed by invoking pgrep without a shell. (2) The open-stub predicate didfn in corpus.stubs(binary)— butstubs()returns addr→Stub, so every card looked "already-banked" and BOTH pools reported ~0 candidates (I nearly concluded the 622-card adapt pile and 954-card weak pile were spent). Fixed to{st.symbol for st in ...values()}and negative-controlled against the known population (35 wave-C cards, 32 banked → exactly 3 still open: PASS). (3) I fired wave D once with hand-typed placeholder card names instead of the generated file — caught and stopped within a minute, relaunched from.run/wave_p31d_args.json. Recorded, not buried (P9). - Pools are NOT spent (post-fix, honest): adapt 569 candidates, weak 890 candidates.
- STRATEGIC (Drew's question "why are our waves so weak?" — he is right): the card lanes I have been drawing from are the mechanical leftovers (12–42-ins cousins), ≈1,400 ins/wave against 612,325 open ins ≈ 0.23%/wave. The Atlas already maps where the mass actually is: head-crack 1,276 groups / 4,248 inst / 186,291 ins (≈3.3 inst per group — a group head crack banks its members), UNKNOWN 1,964 / 138,597, extend-tell 575 / 76,730, redraft 280 / 46,769, jtbl-carve 190 / 45,673. And high-reach groups DO remain: gid
96f8b78ec8d3= 265 instances / 6,084 ins (integration lever, 31-ins exemplar),75cec648ba15= 116 inst / 3,002 ins,152d63c1086d= 64 inst / 1,559 ins. Next: build the wave generator off.run/atlas.jsongroups (top instruction-mass first, routed by the group's measured lever) instead of the leftover card piles.
- THREE instrument defects found+fixed in my own new tooling this cycle (the R32/R35/R39 class, and the reason R39 exists): (1)
-
2026-08-14 (resume session, cont. 2) — THE TARGETING CORRECTION (Drew's "why are our waves so weak?" — he was right). Measured the arithmetic: 635,744 ins open (4.7pp of fleet) across 11,994 fns; the adapt/weak card piles I had been drawing are the 12–42-ins best-seeded tail ⇒ ~1,440 ins/wave = 0.011pp/wave ≈ 440 waves to finish. The mass is elsewhere: cousin-multi 294k ins (avg 44), cold 183k (avg 82, no seed), main-only 38k (avg 87). TWO fixes:
tools/build_wave_atlas.py(NEW) — wave selection off.run/atlas.json, on two measured principles: (a) gate cost scales with (binary,TU) GROUPS, not drafts — each group is a whole-binary rebuild, and wave C was 35 drafts over 27 groups = 1.3 drafts/rebuild (~50 min of gate for 32 banks); the atlas selector concentrates a wave into few binaries (96 drafts → 1 group, ~70× the gate efficiency); (b) mass beats count for the instr-weighted metric. 7,430 draftable candidates available in the agent-lever bands.gate_lanewas STRUCTURALLY BLIND TOmain(R36/R33): it located a stub's home .c byglob('src/<binary>/*.c'), but main's sources live atsrc/*.c→ every main draft grouped undersrc=None. Latent because main has never been wave-gated (main = 79,510 weighted ins at 0.5%, the largest coherent mass left). Fixed to derive fromcorpus.stubs()[..].path; NC'd 3 ways (still-open wave-C drafts 3/3 agree · overlay sample 96/96 agree · main now resolvesNone→src/800.c).
- MAIN PROBE (6 cards, R37 — never spend 96 agents on an unproven path): 4/6 standalone MATCH (67%), including the
-O0boot-module fnfunc_8001099C(the main-specific trap:asm/nonmatchings/boot→src/boot.cis -O0 per §6, so match_one needs--o0; flagged in the mass-lane prompt). Main is agent-draftable. Gate result pending. - RULE VIOLATION CAUGHT (P9, recorded): a wave-E agent wrote its body directly into
src/boot.cinstead of its draft dir — caught bygit status, reverted, gate re-run clean. A dirty tree ABORTS the shared gate for the whole wave, so the wave prompt's HARD RULES were hardened (explicit "never write into the tree, not even to test; no state-changing git commands"). - Wave D (48 adapt cards) drafted 47/48 concurrently — drafting and gating overlap safely (drafts land in
.run/, the gate writessrc/).
-
2026-08-14 (resume session, cont. 3) — MAIN'S BLOCKER DIAGNOSED — it is LINK-LEVEL, not matching. Main probe: 4/6 standalone MATCH (my independent re-verify; the workflow's own repair stage claimed 6/6 — the gate and my check disagree with the agents, R14), then gate 0/4 banked, 4 near. Reproduced one (
func_80013228, a clean 27-insSquare0wrapper) throughharvest_verifyand byte-diffed the built EXE against the original: exactly 2 bytes differ in 413,696, and NOT inside the drafted function — ajalat vaddr 0x80060E74 retargeted fromfunc_80061FA8(game code,build/src/800c2.o) tofirstfile/firstfile2(PsyQ libapi objectbuild/psyq/apicard/A66.o, symbols.us.txt line:firstfile2 = 0x80062248). So adding ONE byte-correct C function tosrc/800.cperturbs symbol resolution between game code and the LINKED PsyQ library objects — the C is right; the link binds a call to a different definition. This is main-specific (main is the only binary withpsyq_integratearchive objects, per theifeq ($(BINARY),main)blocks) and explains why main has sat at 0.5%: it is an INTEGRATION wall, not a matching wall. Main is NOT ready for bulk waves; it needs a named link-resolution investigation lane (candidate leads: duplicate.NON_MATCHINGsymbol definitions visible in the map at bothSquare0andfunc_80061FA8; archive-member selection order when a new undefined ref appears insrc/800.o). Recorded here + decision-log (R31). The 4 main drafts are preserved in.run/wave_p31e/main/as fuel for that lane. -
2026-08-14 — Campaign returned to the lanes that bank. Wave D (48 adapt cards) drafted; independent re-verify 42/48 MATCH (87.5%); 6 DIFF handed to the near/grinder path. Gating now.
Blockers
- main (79,510 weighted ins @ 0.5%) is blocked on a LINK-RESOLUTION defect, not on matching — see the 2-byte
jalretarget above. Needs its own lane before any main wave is worth running. Overlay/md lanes are unaffected and continue to bank.
🛑 SESSION CHECKPOINT (REFRESHED 2026-08-15, overnight campaign mid-flight — do NOT close the phase)
Phase 31 CONTINUES. Overnight campaign running under Drew's "waves and banking all night long" directive (Opus 5, ultracode, 23-min cron heartbeat be8fb48c as the loop's safety net).
This session's arc: resume R22 213/213 → atlas regen (5,144 groups / 11,994 open) → wave-C probe (3 tell) → wave C (35: 11 tell + 24 weak) 32 banked → main probe (6) 0 banked, blocker diagnosed → wave D (48 adapt) 47/48 standalone, gating now. Session banked ≈ 59+ (18 mechanical/probe + 32 wave C + wave D in flight).
THE TWO STRATEGIC FINDINGS (read these first on resume):
- Reach is spent: ~1.25×. 32 wave-C exemplars → only 8 had a sharer, ×2 each. The ×134 era ended in P25/29/30. Fleet-% now moves ~1:1 with functions banked ⇒ throughput is the lever, and the throughput bottleneck is the GATE, whose cost scales with (binary,TU) groups, not drafts (wave C: 35 drafts/27 groups; wave D: 42/23).
tools/build_wave_atlas.py(NEW) concentrates a wave into few binaries (96 drafts → 1 group) and weights by instruction mass — use it for every future wave; the adapt/weak card piles are the 12–42-ins tail (~0.011pp fleet per 48-card wave ≈ 440 waves to finish). - main is BLOCKED on a LINK defect, not on matching (see Blockers).
--exclude-bins mainis the default inbuild_wave_atlas.py.
RESUME STEPS:
pgrep -f tools/gate_lane— never run two gates, and never runbuild_wave*.pyduring one (R35 guard:corpus.stubs()misreports substituted drafts).- Gate the late-repaired wave-D drafts:
.run/wave_p31d_late_slate.json(5 verified MATCH, rescued by the repair stage after the main slate was built). - R22 (
make clean && make extract-all && make check-all→ 213/213), then commit (task + this log together). - Next wave:
.venv/bin/python tools/build_wave_atlas.py .run/wave_p31f_cards.json 96 --max-bins 8→ convert to args ({wavedir,cards_file,cards}) → Workflowscratchpad/p31_wave.js. Lanes:mass(atlas),adapt/weak(tools/build_wave.py <pool>; adapt 569 + weak 890 candidates remain, both verified live after the predicate fix). - Grinder queue has 3 fresh high-value seeds incl.
func_80183578close=1 DELAY-SLOT (§60a precedent: a close=1 delay-slot banked in ~6 min) andfunc_8017DAECcount-exact 113=113.
Watch-fors (all bit tonight): agent self-reports run OPTIMISTIC — always re-verify with match_one yourself, then the gate (wave C claimed 35/35→32 banked; main probe claimed 6/6→I measured 4/6→0 banked). An agent once wrote its body straight into src/boot.c (reverted; prompt hardened) — a dirty tree ABORTS the shared gate. pgrep -f self-matches its own shell wrapper (invoke without shell=True). corpus.stubs() is addr→Stub, not names.
(superseded) checkpoint — 2026-08-14 pre-overnight
Phase 31 CONTINUES (campaign-to-ceiling; do NOT close). Session totals: 65 banked (17 mechanical @$0 + wave A 12+1prop @~117k tok/bank + wave B 35/37 gated, 95% conversion @~88k tok/bank — the decl-matching lesson nearly eliminated integration failures). Stubs ≈ 11,995 (from 12,059). All banks byte-gated + committed; wave banks propagated where sharers existed.
RESUME STEPS (fresh session, after the standard load order):
git log --oneline -20to see the wave-B bank commits; run R22 (make clean && make extract-all && make check-all→ expect 213/213) — it was NOT run after wave B (context ran out; the per-bank gates each verified their own binary, but the standing clean-fleet proof is owed FIRST).make atlas(regenerates maps + atlas post-banks, ~15 min, $0).- Wave C is STAGED, not launched:
.run/wave_p31c_cards.json(14 §172b tell-cards [sonnet] + 24 weak-seed haiku cards — measures the two untested agent lanes). Launch via the persisted workflow scriptworkflows/scripts/p31-adapt-wave-a-wf_2fbef223-859.jspattern (args = the cards; NOTE the tell/weak cards have different fields than adapt cards — adapt the prompt per lane or write a v2 script). Needs/effort ultracode(R27). - Adapt pile remains ~630 SMALL-EDIT cards — the proven 73-95% lane; wave D+ = next 48 by the same selection (exclude banked; see
.run/wave_p31{a,b}_cards.jsonfor taken). - Grinder queue armed: 59 warmstart + 6 wave-A NEARs + 8 wave-B NEARs (4 at close ≤3). Relaunch:
GATE_PHASE=phase-31 .venv/bin/python tools/grinder.py --once --batch 15 …(single-writer: never while a gate runs). - Ledger discipline: velocity row per wave (the table above); distill lessons per R16/R30; close the phase ONLY on measured multi-session yield decay (plan file §Leg-C).
Watch-fors: gate_lane aborts on dirty src/config (clean first); symfix-first before every gate (§173); the safety-classifier can rate-limit under 48-agent bursts (harmless — retry).
(superseded) previous checkpoint (end of build arc)
T0–T9 ALL COMPLETE AND COMMITTED (through commit:2181). Phase totals: 17 banked, 0 agent tokens; stubs 12,059 → 12,042; R22 213/213 verified twice (post-T1, post-T6). The machine: the Atlas (5,139 groups, make atlas), the widened lanes (symfix STALE-DELTA, recover_integration isolation + macro-externs/tu-scope, family_align + len_tells + lenmiss routing), the armed queues (grinder: 59 warmstart records; cards: 954 weak + 192 len + 704 adapt; permuter-49). NEXT = T10+ the campaign loop: L3 grinder running in background (launched at checkpoint time); card/crack WAVES need Drew's /effort ultracode toggle first (R27) — prompt and WAIT. Campaign cadence + close criterion: the plan file §Leg-C. If resuming fresh: read the approved plan + this log; check .run/auto/grinder_heartbeat.json; run make atlas to refresh; continue the loop.