Skip to content

Paper A's revision: answer the external review, advance the pin to v5.12.0, carry W1 - #32

Merged
Hafeok merged 9 commits into
mainfrom
claude/paper-a-w1-revision-b8sx45
Aug 31, 2026
Merged

Hafeok merged 9 commits into
mainfrom
claude/paper-a-w1-revision-b8sx45

Conversation

@Hafeok

@Hafeok Hafeok commented Aug 31, 2026

Copy link
Copy Markdown
Collaborator

Interactive paper revision, five gates, held at every one. Projection work: no claim, term or decision is filed, amended or retired in either repository, except the pin-advance decision this repository owes for its own pin. actor-indexed-determination is untouched.

Session record: meta/sessions/2026-08-30-paper-a-revision/. Manifest: manifest.md there.


What this does

R-1 The related-work survey. Seven works engaged, every locator verified against Crossref, Open Library or NCBI
R-2 The Missing Parameter → Actor-Indexed Determination; the novelty claim narrowed to the review's own formulation; five warrant-overrunning formulations repaired
R-3 Pin v5.9.0 → v5.12.0, predicted before it ran and verified after; four failing quotations repaired, and a fifth repair no checker could see
R-4 The supplement split. The apparatus moved, not reduced, and still running
R-5 W1: 15 moved, 16 deferred, of the 31 S5 occurrences in the two manuscripts
R-6 §7's arity mapped · §8.1's provenance made hypotheses · §8.5's escapes made candidates · §9.5's Study 0 added · the closure ladder repaired

The pin advance held its prediction in every limb

Committed at 7c9e2fa before graph/upstream.yaml was touched. Advanced in two stages — ref first, hashes second — so the firing was observed rather than assumed.

Predicted Observed
E12 0 0
W5 exactly 1 — DDD-frame-15, projected → retired 1, that one
W6 exactly 6, named by id, field and both hashes 6, every id and hash as written
W7 1, unchanged 1

After re-instrumentation: 71 pins, 0 basis-loss, 0 content-drift, 1 standing shadow. Filed as DDD-dec-34.

Three findings that are not repairs

A node forcing the largest repair in this revision fired nothing. DDD-measure-06 retires at v5.12.0 from established. W5 sees only pinned ids and it was not pinned; both existing checkers passed all three prose lines calling it established, because none is a block quotation and none is an appendix row. Found by a sweep written at Gate 1, now shipped as check-status.py, a third checker covering running prose. Four nodes pinned in consequence.

The corroboration check fired on a claim that would have helped the paper. A search result attributed a "Law of Conservation of Complexity" to Woods & Hollnagel's Patterns — which, if true, would have given the framework's own conservation principle a named precedent inside the literature the review says it ignores. It could not be corroborated in any primary source and is used nowhere.

A count in the session's own charter did not reconcile. The charter said 88% of W1 lives in the two papers. Against the committed classification, 88 is the corpus-wide mutable total and the two manuscripts hold 31 of it — 35%. The migration seed is corrected; W1's remaining surface is 73 of 88, not zero.

The survey corrected the paper, and corrected the review

The review's own Hollnagel-and-Woods locator (10.1201/9781420005684) resolves to Woods & Hollnagel (2006), Patterns, while its prose concerns Hollnagel & Woods (2005), Foundations. Both are now cited; the correction goes to the response-to-review rather than being made silently.

Three entries record where a neighbour is stronger than the framework — Horvitz supplies a decision criterion the framework lacks off the closing region; Leveson carries authority structurally; and four of DDD-frame-08's five accountability elements are already Bovens's, leaving authority linkage and a change of tense. One element and a tense, not a new theory.

One difference claim was withdrawn by the reading itself. The Gate 1 plan said Matthias's responsibility gap is "fixable rather than novel". The primary text does not support fixable, and the paper says so in its own §11.

The full read caught two defects all three checkers missed

The abstract still carried the retired four-mode enumeration after §4.1 had been repaired — the paper's most-read paragraph projecting a claim retired at v5.10. And two references to "Appendix A" survived in a paper whose appendix had moved. Both fixed, and both recorded: three green checkers and a clean validator run saw neither.

Length, with the method and the attribution

Body 11,708 → 17,334 words (+48%), supplement 5,232, by the method that reproduces the review's own figures to within 1%. No argument was cut. The machinery left the argumentative path; what replaced it is qualification — a narrower claim needs the hedging a broad one does not, and three survey entries conceding a neighbour's strength each cost a paragraph.

Close checks

Gate Result
check-quotations.py @ v5.12.0 30 verbatim, 0 disclosed-partial, 0 failing
check-status.py @ v5.12.0 96 inline status assertions, 0 wrong
check-appendix.py @ v5.12.0 75 nodes rendered, 75 cited, 0 discrepancies
gen-appendix.py idempotence byte-identical across three runs, manuscript unwritten
validate-core-order.py core/ 0 errors, 0 warnings; 71 pins, 0/0/1
validate-claims.py claims / decisions valid; 26 claims with the 6 recorded warnings; 22 decisions
reference closure 22 entries, 0 cited-but-absent, 0 absent-but-cited
cross-references 58 headings, 0 dangling § references

Not in scope, and not bundled

Filing anything in canon, including Q44 and Q45, which were read and cited nowhere. The migration proper. The primer. The transfer. W4. The decoder repair. Six items are booked at meta/successor-items-paper-a-revision.md; the response-to-review's §8 lists what the paper still owes, and the first item on it is the measure's construct-validity repair, which this revision does not make.

🤖 Generated with Claude Code

https://claude.ai/code/session_011Cahdg8Re8TTjvheiHLNrA


Generated by Claude Code

Hafeok added 9 commits August 30, 2026 18:53
The session's first act, before any repository is read for the work:
prompt.md and bootstrap.md land under
meta/sessions/2026-08-30-paper-a-revision/, with the package README
under inputs/ and the sessions index extended.

Five of the seven package files are already committed byte-identically
— the review and the triage at 2026-08-23-phase1a/, and Q44, Q45 and
the ontology explainer at 2026-08-27-ground-migration/inputs/. Each was
hashed against its committed copy before the decision not to re-file,
and those paths are the ones this session cites. Filing Q44 and Q45
nowhere new is deliberate: they are unfiled downstream demand, the
prompt keeps them unfiled, and the paper may not cite them.

Base commits recorded in bootstrap.md: actor-indexed-determination at
81f6929, which annotated tag v5.12.0 resolves to exactly; this
repository at 54f00eb.

Basis: DDD-dec-20; DDD-dec-17; DDD-delivery-02; term:delivery
…, three plans

Nothing in papers/ is touched. No claim, term or decision is filed, amended
or retired in either repository, and the pin stays at v5.9.0. The whole of
this diff is the gate report and its two instruments.

CITATIONS. All 72 nodes Paper A cites still resolve at v5.12.0 — no E12-class
loss. check-quotations.py reports exactly the four failing quotations the
prompt predicted (DDD-frame-15's statement, DDD-frame-02's, term:residual-
discretion's, and the compact form at L1208, whose citation moves to
DDD-frame-17 while its prose does not). check-appendix.py reports ten
discrepancies across eight nodes, all content movement.

A FIFTH REPAIR NO CHECKER SEES. DDD-measure-06 is retired at v5.12.0,
superseded by DDD-measure-16 and DDD-measure-17 — the review's F-A, already
repaired in canon. The paper asserts it inline as **established** three
times and neither existing checker reads an inline status assertion.
status-sweep.py closes that gap: 68 assertions checked, 0 wrong at v5.9.0,
4 wrong at v5.12.0.

W1. 18 proposed moved, 13 deferred, of the 31 S5 occurrences in the two
manuscripts. The deferrals are enumerated line by line with a reason each:
generated appendix rows, passages quoting a live claim verbatim, and one
filename. The ambiguous "ground the task faces" pair is separated and
flagged rather than ruled, per the migration's own constraint.

A COUNT IN THE CHARTER THAT DOES NOT RECONCILE. 88 is the corpus-wide
mutable S5 total, not the count in the two papers. They hold 31 of it —
35%, not 88%. Reported rather than worked around; the close report will
say 31 of 88 and the migration's remaining W1 surface is 57.

Basis: DDD-dec-20; DDD-agent-01; DDD-delivery-02; term:escape; DDD-frame-15;
DDD-frame-17; DDD-measure-06; DDD-measure-16; DDD-measure-17; DDD-frame-02;
term:residual-discretion; DDD-measure-12; term:verdict
…fted

R-1 drafted. §11 is rewritten: seven inherited neighbourhoods compressed and
kept, the seven works the review named added with what the framework takes
and where it differs, a comparison table, and the narrowed novelty claim.
The old §11.1 becomes §11.6 unchanged.

READING BEFORE DRAFTING. gate2-survey-notes.md is the reading record: how
each locator was verified (Crossref, Open Library, NCBI — no publisher page
was reachable and none is the basis of any entry), and at what grade each
text was read. Four works were read primary and in full; the two Hollnagel
and Woods volumes were not obtained, and both the notes and the reference
entries say so. Verifying a locator is not reading a text, and the
References preamble now marks the two separately.

TWO CATCHES. The review's own locator for Hollnagel and Woods resolves to
Woods & Hollnagel (2006), PATTERNS, while its prose concerns Hollnagel &
Woods (2005), FOUNDATIONS — two books, reversed author order, one shared
main title. Both are cited; the error goes to the response-to-review rather
than being fixed quietly, on Emil's ruling. And a "law of conservation of
complexity" attributed by a search result to Patterns could not be
corroborated and is used nowhere; it is written down so a later session does
not rediscover and believe it.

THE SURVEY CORRECTED A DIFFERENCE CLAIM. The Gate 1 plan said Matthias's
responsibility gap reads as "an arrangement naming an executor and no
principal — fixable rather than novel". The reading does not support
"fixable": the gap rests on a control condition on just ascription, and
whether structural completeness answers it is a normative question
DDD-frame-08 is projected about. The claim is withdrawn in the section
itself rather than carried quietly. Three entries now record where a
neighbour is stronger than the framework, Horvitz's decision criterion
first among them.

The section adds no node to Appendix A and no pin: all three checkers are
green at the pin — 29 verbatim / 0 failing, 72 nodes / 0 discrepancies,
75 inline status assertions / 0 wrong.

ALSO LANDED, on Gate 1 rulings: the stale-line-number finding added to the
migration seed's method rule ("classification data cites content, never
position"), with W1's counts corrected there from 29 to 31 in the two
manuscripts and the corpus-wide 88 restated; and the ambiguous set's five
occurrences moved to defer in w1-enumerate.py, which now reports 15 moved,
16 deferred.

LENGTH, honestly. Body 11,708 -> 14,069 words; total 15,250 -> 17,611;
headings 63 -> 68. That is over the 1,200-1,600 estimated at Gate 1, and
the overrun is qualification: the per-entry reading grade, the comparison
table, and the three places the framework comes off worse all cost words.

Basis: DDD-frame-01; DDD-frame-03; DDD-frame-05; DDD-frame-08; DDD-frame-13;
DDD-cost-09; DDD-delivery-01; DDD-delivery-02; DDD-measure-01;
term:arrangement; term:accountability; term:commitment-level; DDD-dec-20
…hetoric

pass, and the review's small items

R-2. Retitled to candidate 1 on Emil's Gate 1 ruling: "Actor-Indexed
Determination", with the existing subtitle promoted unchanged. The narrowed
claim is stated once, in §1.2, in the review's own terms — a specific,
auditable synthesis of resolution, assurance, delivery and accountability
over the arrangement as the unit — together with the withdrawal it rests on:
the arrangement is not this framework's, the absence the old title asserted
does not hold, and §11 credits whose it is. The abstract and §12 carry the
same narrowing.

RHETORIC. "Two results give the framework its shape" becomes two proposals
with declared falsifiers. "Its strongest result" (§6, §12) and "the
strongest defensible result" (§5.4) become load-bearing proposals, each
labelled projected at the point of use. §6.2's "narrower result" is labelled
reported and the difference from observed is stated. The conclusion's
established list now says every claim in it is formal, and that established
means internally argued and unchallenged, not externally validated. The
closing paragraph no longer asserts the framework's value; it states its aim
and names the open question.

R-6.2 Accountability counts reconciled in §7 with a mapping table. The
five-element claim refines term:accountability's third element into stake
and sanction path and ADDS authority linkage. A projected claim adding an
element to a settled term is a supersession question; the paper flags that
canon disagrees with itself about the relation's arity and does not take the
ruling.

R-6.3 §8.1's provenance assignments become hypotheses with what would
overturn each, and the schema row — the one the review challenged — is shown
to sit under two provenances at once, which the five-way partition does not
adjudicate. §8.5's two escapes become CANDIDATE escapes: an escape is a
property of an arrangement, so a walk that did not look everywhere has not
established one, and code search, contract tests, telemetry or asking the
consuming teams all reach the question the walk assumed missing.

R-6.1 §9.5 gains Study 0 — coding reliability first, per dimension and not
pooled — ahead of any comparative work. Two admissions follow: the
escaped-decision count is circular until the instrument holds, and failure
of the umbrella prediction would not falsify the ontology, because the
framework's descriptive and predictive claims are separately falsifiable.

NOT IN THE PROMPT'S R-6 LIST, delivered and flagged for ruling: the closure
ladder's axis mixing, which the triage records as Emil's own Gate 1 ruling
at the drafting session. The ladder now ends at constructively closed and is
three rungs, all operational; decidability is stated as the logical axis's
answer with both directions of non-implication given. Revert if this exceeds
the gate.

Three stale citations of the old title updated: the reviewer brief, the
measure note's paper context, and meta/way-of-working.md. README.md:32 and
CHANGELOG.md:303 use the phrase rather than the title and are left alone.

All instruments green at the pin: 29 verbatim / 0 failing, 72 nodes / 0
discrepancies, 85 inline status assertions / 0 wrong; the three repository
validators pass with the six recorded warnings.

Basis: DDD-frame-01; DDD-frame-04; DDD-frame-05; DDD-frame-07; DDD-frame-08;
DDD-frame-11; DDD-floor-02; DDD-hyp-04; DDD-dec-26; term:closure;
term:accountability; term:attribution; term:answerability; term:liability;
term:commitment-level
…moves

Committed ahead of the operation so the advance is observed rather than
assumed, per the DDD-dec-29 pattern. graph/upstream.yaml still reads
ref: v5.9.0 at this commit.

Predicted: E12 0 · W5 exactly 1 (DDD-frame-15, projected -> retired) ·
W6 exactly 6 (DDD-cost-09 region; DDD-delivery-01, DDD-frame-02 statement;
DDD-frame-15 statement and region; term:delivery and
term:residual-discretion canonical_md) · W7 1, unchanged. 61 of 67 pins
still.

The three the migration seed predicted for this advance — term:delivery,
DDD-cost-09, DDD-delivery-01 — appear exactly. The other three are Phase
1a's repairs, which the seed does not list because they are not
ground-migration work. None of the seed's seven W1 W6 fires here, and if
one does the prediction is wrong and the advance stops.

DDD-measure-06 retires at v5.12.0 and fires NOTHING, because it is not
pinned — the node forcing the largest prose repair in this revision is
invisible to the pin instrument. Pinning it and its two successors is
proposed at the advance as a ruling, not a mechanical step.

Basis: DDD-dec-29; DDD-dec-28; DDD-agent-01; DDD-frame-15; DDD-frame-02;
DDD-cost-09; DDD-delivery-01; term:delivery; term:residual-discretion
repairs, Appendix A regenerated, the supplement split, and W1

THE ADVANCE HELD ITS PREDICTION IN EVERY LIMB. Predicted and committed at
7c9e2fa, before graph/upstream.yaml moved: E12 0, W5 exactly 1, W6 exactly
6, W7 1 unchanged. Observed with ref advanced and hashes deliberately NOT
re-instrumented: "67 pins resolved, 1 basis-loss, 6 content-drift, 1
shadowed id" — every id and every hash the one written down beforehand.
After re-instrumentation: 71 pins, 0/0/1. The advance is DDD-dec-34.

FOUR PINS ADDED, and it is a finding rather than housekeeping.
DDD-measure-06 retires at v5.12.0 from established, forces the largest
repair in this revision, and FIRED NOTHING — W5 only sees ids that are
pinned. DDD-measure-06, DDD-measure-16, DDD-measure-17 and DDD-frame-17
are now pinned.

THE FOUR QUOTATIONS, all repaired: DDD-frame-15's statement (retired; §4.1
now carries DDD-frame-17's three values, and says the review killed the
four modes), DDD-frame-02's amended clause, term:residual-discretion's
amended text with the cryptographic-hash separation, and the compact form
— the cheapest, exactly as predicted: it survives verbatim in
DDD-frame-17's closing clause, so the citation moved and no prose did.

AND THE FIFTH, which no checker saw. DDD-measure-06's biconditional is
replaced at six sites by the three separated conditions — existence,
availability, estimability — with closure governing only availability.
§4.4 records the provenance: the measure note had already conceded both
failure directions, canon was never updated, so this projection was
stronger than the note sharing its graph, and an external review found it
rather than an instrument.

APPENDIX A regenerated wholesale against the tag, never hand-edited: 72
nodes and 13 discrepancies before, 75 nodes and 0 after, idempotent across
three runs byte-compared. The independent re-read confirms it.

THE SUPPLEMENT SPLIT on the ruled boundary — about the graph moves, about
the world or about warrant stays. paper-a-supplement.md carries the
reproduced nodes, the pin, the filing statuses, the arity finding, the W1
record and Reproduction. Every projected-and-empty sentence stayed in the
paper. gen-appendix.py and check-appendix.py learned a fourth argument so
citations are read from the manuscript and rendered into the supplement;
both carry the new defect history, including the failure mode where
passing the supplement as the manuscript renders an empty appendix and
exits 0.

check-status.py ships as the third checker (GATE 2 ruling), with its own
defect history and its origin recorded as its warrant.

W1: 15 moved, 16 deferred, of the 31 in the two manuscripts. Every
deferral verified untouched by explicit fragment check. §5.4's heading
renames with one clause naming the asset it now diverges from — flagged,
not ruled.

Checkers at v5.12.0: 30 verbatim / 0 failing · 93 status assertions / 0
wrong · 75 nodes / 0 discrepancies. Repository gates green with the six
recorded warnings.

Basis: DDD-dec-34; DDD-dec-29; DDD-dec-28; DDD-agent-01; DDD-frame-02;
DDD-frame-15; DDD-frame-17; DDD-measure-06; DDD-measure-16; DDD-measure-17;
DDD-measure-12; term:residual-discretion; term:delivery; term:verdict;
term:accountability; DDD-frame-08; DDD-dec-26
… commit

The supplement's S2 carried a placeholder while the note's W1 edits were
uncommitted — a hash cannot be cited before it exists. Now resolved to the
commit this session created, with the reason stated: Paper A pinned the note
at aa7e135, W1 edited the note, so the pin advances to a commit this revision
produced rather than citing prose that says ground distribution while this
paper says deployment distribution.

Basis: DDD-dec-34; DDD-delivery-02; term:delivery
read caught

THE FULL READ CAUGHT TWO REAL DEFECTS, both invisible to all three
checkers because neither is a block quotation, an appendix row, or a
status assertion.

  1. The ABSTRACT still carried the retired four-mode enumeration —
     "by a filed decision, an actor's judgment, an arrangement default, or
     an uncontrolled draw" — after §4.1 had been repaired to
     DDD-frame-17's three values. The paper's most-read paragraph was
     projecting a claim retired at v5.10. Fixed.
  2. Two references to "Appendix A" survived in a paper whose appendix had
     moved to the supplement. Fixed.

Recorded rather than quietly corrected because they are the argument for
the full read: three green checkers and a clean validator run did not see
either, and the second checker gap this session found is the same shape as
the first.

ALSO from the read: §3.3's trained-model case now works the review's own
counterexample explicitly — policy-committed on one axis, predetermined on
another, standing on a third, all true at once and none competing. The old
list read as competition because it drew values from three axes at once.
And the abstract's opening names WHICH accounts held the determiner fixed
rather than implying all of them did, which the survey no longer supports.

response-to-review.md maps all twenty-two objections: twenty-one conceded,
repaired, or accepted and booked; ONE defended, a wording choice whose
underlying finding is conceded. The headline was corrected during drafting
— an earlier version claimed three defences by counting propositions the
review PRAISED, which would have flattered us.

The manifest records the close checks, the word count with its method and
the growth attributed, the pin advance's prediction against its
observation, the four gates' rulings, and three findings that are not
repairs: the unpinned node that fired nothing, the corroboration check
firing on a claim that would have HELPED the paper, and the charter's 88%
that was really 35%.

Successor item 4 carries Emil's GATE 4 rule: an advance without
instruments is unobservable, and an unobservable advance is presumed
discharge at the pin layer.

Close checks, all green: 30 verbatim / 0 failing · 96 status assertions /
0 wrong · 75 nodes / 0 discrepancies · validate-core-order 0 errors 0
warnings, 71 pins 0/0/1 · 26 claims with the 6 recorded warnings · 22
decisions · 22 references with 0 cited-but-absent and 0 absent-but-cited ·
58 headings with 0 dangling section references.

Body 11,708 -> 17,334 words (+48%), supplement 5,232, by the method that
reproduces the review's own figures to within 1%.

actor-indexed-determination is untouched, which is the correct state for a
projection session.

Basis: DDD-dec-20; DDD-dec-34; DDD-agent-01; DDD-frame-15; DDD-frame-17;
DDD-frame-16; DDD-frame-08; DDD-measure-06; DDD-measure-16; DDD-cost-20;
term:commitment-level; term:accountability; term:residual-discretion;
term:delivery
…form

On Emil's GATE 5 ruling: record the abstract finding in the manifest as the
argument for the full read, and state the general form because it will
recur.

THE FORM. Instruments cover the surfaces someone thought to instrument. The
abstract is the surface nobody thinks to instrument, because it is prose
ABOUT the paper rather than a citation IN it — so it inherits no citation's
protection while carrying more of the reader's warrant than any single
citation does.

This was the second unwatched-surface finding of the session and the same
shape as the first: §5.1's was an unpinned node moving silently, this one
is unciting prose asserting silently. A third checker answered the first. A
fourth is deliberately NOT proposed for this one — an instrument that
parsed an abstract for claims it does not cite would guess at meaning
rather than check correspondence, and its false negatives would be
indistinguishable from a clean run. The three checkers verify
correspondence; that is what makes them trustworthy and what bounds them.
The remedy is the full read, and its cost is now known.

Successor item 7 carries the standing instruction: in any session
repairing a projection against a moved graph, the abstract and the
conclusion are repaired last and read whole, because a restatement
inherits nothing from the citation it paraphrases. Whether it belongs in
meta/way-of-working.md as a projection rule is left to whoever next
touches that document, and is not filed here.

Manifest §5.2 also now records both checks that fired against claims
which would have HELPED the paper — the unverifiable conservation law, and
this session's own draft headline claiming three defences by counting
propositions the review praised. Twice in one session, which is the only
evidence that the checks are not decorative.

Basis: DDD-dec-20; DDD-agent-01; DDD-delivery-02; DDD-frame-15;
DDD-frame-17; term:escape
@Hafeok
Hafeok merged commit d89ed55 into main Aug 31, 2026
1 check passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant