Files
LithosAnanake/experiments/std79-doe/README.md
T
Robert Allan JamesandClaude Sonnet 5 e51a8d229e
Build / build-amd64-iso (push) Waiting to run
Build / build-aarch64-iso (push) Waiting to run
Build / build-riscv64-img (push) Waiting to run
Fix stadium_grant_quota() donor floor; rerun std79 DoE clean, 81/81 (FABRIC-3.md §XVII)
capsule_birth.c hardcoded every new VM's initial Stadium quota grant to split
from Hera specifically. Since a grant always halves whatever the donor
currently has, Hera's own free list converges toward empty after a bounded
number of grants — independent of whether the Stadium as a whole still had
spare capacity, since VMs she'd granted to earlier typically still held
nearly all of their own share untouched. Past that point every subsequent
VM birth's Stadium grant would be silently refused (soft-failed, non-fatal
by existing design), even with plenty of capacity sitting idle elsewhere.

Fixed by adding an O(1)-maintained free_count to StadiumVMQuota (incremented
in stadium_evict(), decremented at both of stadium_admit()'s free-list-pop
sites, set/adjusted in stadium_grant_quota()'s own split — this also let
grant_quota drop its old O(free-list length) counting walk in favor of an
O(1) read) and stadium_best_donor(), an O(live VM count) scan over quota
slots returning whichever in-use VM currently has the most free cells.
capsule_birth.c's birth path now splits from that VM instead of
unconditionally vm_uuid_hera().

Verified with another full rerun of the 3x9x3 std79 DoE campaign from
scratch — same discipline as the prior Stadium fix (any defect repair
reruns the whole DoE from the top) — one continuous boot per architecture,
all 9 identities simultaneously live throughout. 81/81 trials correct, 0
mismatches, DOE-RUN header sequence md5-identical to every prior run.
aarch64 ~280s total (vs ~290s for the O(ncells)-scan fix alone — confirms
no regression). Both known Stadium defects are now closed together on one
clean campaign rerun.

Co-Authored-By: Claude Sonnet 5 <noreply@anthropic.com>
Claude-Session: https://claude.ai/code/session_01EXieurDfDSsDFdnSyusuWo
2026-09-11 17:20:07 -04:00

3.2 KiB
Raw Blame History

std79 DoE — 3(architecture) × 9(identity) × 3(replicate) randomized full-factorial

The formal successor to experiments/std79-exerciser/'s ad hoc campaign (see FABRIC-3.md §XII), requested as a genuine randomized full-factorial design matching this project's own DoE methodology (capsules/doe.4th's Fisher-Yates run-matrix shuffle) rather than convenience batching — and written entirely in FORTH, not host-orchestrated shell scripting. See FABRIC-3.md §XV for the design writeup, §XVI for the real aarch64 Stadium/COOL scaling bug that was found, root-caused, and fixed (src/starkernel/vm/stadium.c), and §XVII for a second, related Stadium bug (stadium_grant_quota()'s donor-floor) found while writing up §XVI and fixed in a follow-up pass — the 3-boot batching workaround below is now historical only; a single boot with all 9 identities simultaneously live works cleanly on all three architectures as of both fixes.

std79-doe.fth is not a capsule loaded via EXEC — feed it as raw text to a running REPL (e.g. socat - UNIX-CONNECT:<serial_sock> < std79-doe.fth), same as the original exerciser, then invoke EXEC-STD79-DOE ( seed lo hi -- ) once loaded. lo/hi select which identity-index range (0-8) this boot's live VMs cover — pass 0 8 for a single boot with all 9 identities simultaneously attached (now confirmed working on all three architectures, see §XVI); a narrower range (e.g. 0 2, 3 5, 6 8) still works too, for running the DoE across multiple smaller boots if ever needed for an unrelated reason. Every boot must use the same seed so INIT-MATRIX/SHUFFLE-MATRIX reproduce the identical 27-slot master permutation; only the lo/hi filter differs, so run_id (always the slot's true position in the master shuffle, 0-26) stays directly comparable across boots.

results-20260911/ holds the raw captured serial output from the original campaign (pre-fix): amd64-doe-raw.log (all 27 trials, one boot, all 9 identities simultaneously live) and {aarch64,riscv64}-doe-batch{1,2,3}-raw.log (27 trials each, split across 3 boots of 3 identities, the then-necessary workaround). Result: 81/81 trials correct, 0 mismatches.

results-20260911-stadium-fix/ holds the rerun after §XVI's O(ncells)-scan fix — one boot per architecture, all 9 identities simultaneously live in every boot, same seed throughout. Result: 81/81 trials correct, 0 mismatches, byte-identical output to the original campaign — and the DOE-RUN header sequence (run_id/id_idx/id_label/rep assignment) is md5-identical across all three raw logs, confirming the master shuffle is genuinely architecture-independent. Total per-architecture wall-clock (boot + all 9 identity attaches + full 27-trial run, one continuous QEMU session): amd64 ~165s, aarch64 ~290s, riscv64 ~161s — aarch64's identity 04 attach alone, which stalled 90+ minutes before the fix, now completes in ~34s.

results-20260911-donor-floor-fix/ holds a further rerun after §XVII's donor-floor fix (same discipline: any Stadium defect repair reruns the whole DoE from the top). Result: 81/81 trials correct, 0 mismatches, DOE-RUN header sequence md5-identical to every prior run. Total wall-clock: amd64 ~161s, aarch64 ~280s (confirms no regression from §XVI's fix), riscv64 ~162s.