Skip to content

docs(reprint): the ledger carries four runs and the cost is 4736 s - #323

Open
the-homeless-god wants to merge 3 commits into
devfrom
a/user-stories-and-the-reprint-ledger
Open

the-homeless-god wants to merge 3 commits into
devfrom
a/user-stories-and-the-reprint-ledger

Conversation

@the-homeless-god

@the-homeless-god the-homeless-god commented Oct 4, 2026 •

Copy link
Copy Markdown
Member

Task 8570 claimed the reprint takes ten hours against a goal of two. The claim
was false by a number, and the number was a gap in the ledger: 36839 s is the
last RECORDED row of docs/reprint-ledger.tsv (17 September 2026), not the
cost. Runs after it were many and none was written down.

Four rows written in

date what outcome seconds peak GiB bytes cap source
2026-09-25 print landed 128700 0 0 4e12 header of scripts/bootstrap-reprint.sh (35 h 45 min); peak and bytes never recorded, no log named in the tree (task 6704)
2026-10-04 print landed 4736 50.6 44661951 4e12 gpu, /home/u/b-print-0725-print.log, read only: start 18:24:13 UTC from the first line, end 19:43:09 UTC from stat -c %y
2026-10-04 print killed 4989 41.6 0 4e12 dev, /home/b/projects/flang-cluster/print-batch.log: SIGTERM from the harness at 19 % of the judging stage
2026-10-04 print landed 6890 0 0 0 WITHOUT THE JUDGE, 7 files, rebuild 1 min 17 s. FOREIGN measurement: came in an agent report, no log named, not measured here

The landed run of 4 October printed 7 files and 44661951 bytes, and its judging
chain ate 2508501226319 steps of 4000000000000 (62 %) - the print said that
itself. The copy had no git, so the fingerprint says commit no-git; there is
no hash for the run.

Two warnings added to the ledger header

  1. Rows before 1 October are not comparable with today's. They were taken
    by a binary WITHOUT the kernel call memo (ADR-0051, task 5709). 36839 s
    against 4736 s is 7.8 times, and that is the memo, not the tree.
  2. Steps per second is not a measure of speed. A memoized call is charged
    its steps without being executed, so between two neighbouring pulses of one
    and the same print the counter ran 85 million and then 2.4 billion steps per
    second. Cost is written as seconds of a landed run; a percentage of the cap
    is spend, not speed.

Task 8570 rewritten and renamed

8570-the-print-is-outrun-because-it-is-slow.md ->
8570-the-print-runs-out-of-steps-not-of-hours.md. The goal of two hours is
met (point 3 of its own acceptance: a row under 7200 s). What is left is named:
the step cap, 62 % of which the judging stage ate, raised from 4 to 8 trillion
by 3f9b9a77b (on dev it is not yet). The instrument that measures the cost by
itself is named too: the print's own warning at the half-cap threshold,
flang/src/emit/c/flang_repl.c:2324. The three places that convert runs to one
machine through steps per second are named with line numbers and called an
unfit measure; they are not edited here.

scripts/seed/reprint-freshness.fscript quoted this ledger as "from 10 h 14 min
to 35 h 45 min, peak 44.9 GiB" and now says from 1 h 18 min, peak 50.6 GiB.

The script header too, with zero growth of lines

A stale number there changed someone's work on 4 October: an executor read in the
header of scripts/bootstrap-reprint.sh that a landed run costs between
6 h 23 min and 35 h 45 min and did not run the reprint. The growth gate
refuses added lines in shell, so 14 lines are replaced by 14 and the file is
still 3592 lines (bash -n and sh -n clean):

place was is
the price of --check 8 h 34 min 02 s, 30 August 1 h 18 min 56 s (4736 s), 4 October
the refusal of --bystro "the price is 8 h 34 min" "1 h 19 min"
the print directory "39 MB and half a day" "45 MB and an hour and a bit", 44661951 bytes, peak 50.6 GiB
"the print takes six hours", 5 places six hours an hour and a bit

Not touched on purpose: the MEASURED_COST / MAX_STEPS block, where
"35 h 45 min" also stands. Commit 3f9b9a77b rewrites those very seven lines to
raise the cap, and two edits of one line would collide in the merge. Task 8570
says the stale number stays there until that commit lands.

The second print of 4 October, already under the raised cap of eight trillion,
prints no "step budget running out" warning at all - the raise did what it
was for. Its seconds are not in the ledger: the run is still going, and a price
must not be written from a ceiling.

Checked

All 21 pre-push guards green in 63 s (task-ledger, prose-numbers-guard
224/224, no-growth, lint-growth, tree-inventory among them). flang check
of reprint-freshness.fscript clean; the freshness plan still reads the ledger
and still reports the last byte-for-byte check of 8 September, which these rows
do not touch.

NOT MEASURED: the reprint itself was not run from this worktree - a print is
running on gpu and a second one is not started beside it; and
sh scripts/bootstrap-reprint.sh --bystro was not called for the same reason.
flang/scripts/tasks.fscript --plan 'task board whole' exhausts its own step
limit on the untouched tree as well, so it proves nothing either way.

Task 8570 said the print takes ten hours against a goal of two. The number
came from the last RECORDED row of docs/reprint-ledger.tsv, 17 September
2026, 36839 s. Runs after it were many and not one was written down: the
ledger is filled by hand and has no instrument of completeness (task 6704).
Four missing runs are written in now.

2026-10-04, print, landed, 4736 s (1 h 18 min 56 s), on gpu. Start
18:24:13 UTC from the first line of /home/u/b-print-0725-print.log, end
19:43:09 UTC from the file timestamp (stat -c %y); the log was read, not
touched. Output 7 files, 44661951 bytes, from the line that says so. Peak
50.6 GiB, the largest of 143 pulses of the print itself, not /usr/bin/time;
in the judging stage the peak was 50.0 GiB. The judging chain took
2508501226319 steps of 4000000000000, 62 percent, and the print said that
itself: it prints a warning when a stage passes half the cap. The copy had
no git, so the fingerprint says the seed was printed from commit "no-git" -
there is no hash for this run.

2026-10-04, print, killed, 4989 s (1 h 23 min 09 s), on dev. SIGTERM from
the harness at 19 percent of the judging stage, peak 41.6 GiB. Start
16:53:13 UTC from the log header of
/home/b/projects/flang-cluster/print-batch.log, end from the timestamp of
its last line.

2026-10-04, print WITHOUT THE JUDGE, landed, 6890 s (114 min 50 s), 7 files,
rebuild of the printed sources 1 min 17 s. A FOREIGN measurement: the number
came in an agent report of the same day, no log was named, nothing here
measured it. Peak, bytes and cap are zeroes with the reason spelled out.

2026-09-25, print, landed, 128700 s (35 h 45 min), from the header of
scripts/bootstrap-reprint.sh and .github/workflows/reprint.yml. Peak and
bytes were never recorded and the tree names no log for that run - the same
case as the row of 30 August (task 6704).

The ledger header now carries two warnings. Rows BEFORE 1 October were taken
by a binary WITHOUT the kernel call memo (ADR-0051, task 5709) and are not
comparable: 36839 s against 4736 s is 7.8 times, and that is the memo, not
the tree. And steps per second is NOT a measure of speed - a memoized call
is charged its steps without being executed, so between two neighbouring
pulses of one and the same print the counter ran 85 million and then 2.4
billion steps per second. Cost is written as seconds of a landed run; a
percentage of the cap is spend, not speed.

Task 8570 is rewritten to that and renamed. The goal of two hours is met,
point 3 of its acceptance is done, and what is left is named: the step cap,
which the judging stage ate 62 percent of. It is raised from 4 to 8 trillion
by commit 3f9b9a77b, which on 4 October lives in a branch and is not on dev.
Points 1, 2 and 4 are marked not done or not measured, with the reason. The
three places that convert runs to one machine through steps per second are
named with line numbers and called an unfit measure; they are not edited
here.

scripts/seed/reprint-freshness.fscript told the human the cost was "from
10 h 14 min to 35 h 45 min, peak 44.9 GiB", quoting this ledger. It now says
from 1 h 18 min to 35 h 45 min, peak 50.6 GiB.

Checked: prose-numbers-guard --plan Check, 224 of 224 marks agree with the
tree; task-numbers-guard --plan Check, no doubled numbers; flang check of
reprint-freshness.fscript clean; the freshness plan still reads the ledger
and still reports the last byte-for-byte check of 8 September, which my rows
do not touch. flang/scripts/tasks.fscript --plan "task board whole"
exhausts its own step limit on the untouched tree as well, so it proves
nothing either way.

NOT MEASURED: the reprint itself was not run from this worktree - a print is
running on gpu and a second one on the same machine is not started beside
it; and sh scripts/bootstrap-reprint.sh --bystro was not called for the same
reason.
A stale number in the header of scripts/bootstrap-reprint.sh changed
someone's work on 4 October 2026: an executor read there that a landed run
costs between 6 h 23 min and 35 h 45 min and did NOT run the reprint. A
second one read "8 h 34 min" from docs/reprint-ledger.tsv. That is what a
stale number costs - not a blemish in prose but work cancelled.

The measured cost of the landed run of 4 October is 4736 s, 1 h 18 min 56 s
(gpu, /home/u/b-print-0725-print.log, 18:24:13 to 19:43:09, 7 files and
44661951 bytes out, peak 50.6 GiB over the pulses of the print itself).

The header is rewritten with ZERO growth of lines, because the growth gate
refuses added lines in shell: 14 lines replaced by 14, the file still 3592
lines. What the script now says:

  the price of --check       8 h 34 min 02 s, 30 August -> 1 h 18 min 56 s
                             (4736 s), 4 October
  the refusal of --bystro    "the price is 8 h 34 min" -> "1 h 19 min"
  the print directory        "39 MB and half a day" -> "45 MB and an hour
                             and a bit", 44661951 bytes, peak 50.6 GiB
  "the print takes six       six hours -> an hour and a bit
  hours", five places

Both shells parse it: bash -n and sh -n are clean.

NOT TOUCHED ON PURPOSE: the MEASURED_COST and MAX_STEPS block, where
"35 h 45 min" also stands. Commit 3f9b9a77b rewrites those very seven lines
to raise the step cap, and two edits of one line would collide in the merge.
Task 8570 says so, and says that the stale number stays there until that
commit lands.

Task 8570 also gains two things measured today. The second print of the day,
running under the raised cap of eight trillion, prints NO "step budget
running out" warning at all - the raise did exactly what it was for; its
seconds are not written into the ledger because the run is still going and a
price must not be written from a ceiling. And the three places that convert
runs to one machine through steps per second are named with a reason: the
conversion is an unfit measure, since a memoized call is charged its steps
without being executed.
An awk over docs/reprint-ledger.tsv for the last "print landed" row gave
6890 s - the run WITHOUT the judge, a foreign measurement whose time of day
is not known. The refusal of scripts/bootstrap-reprint.sh --bystro quotes
exactly that row as "the price of the last landed run", so it would have
named the price of a print that never judged anything.

The three rows of 4 October are ordered so the run without the judge stands
first (no time of day to place it by) and the two runs with known starts
keep their order: 16:53:13 UTC killed, 18:24:13 UTC landed. The last landed
row is now 4736 s, the number the script header and the task name. Task 8570
says why the order is what it is.

This branch has not been deployed

No deployments
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant