Alignment machinery design: production-mode reconciliation with displacement disclosure (draft for referee)#217
Alignment machinery design: production-mode reconciliation with displacement disclosure (draft for referee)#217MaxGhenis wants to merge 14 commits into
Conversation
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
|
The latest updates on your projects. Learn more about Vercel for GitHub.
|
Adversarial design referee — alignment machinery (draft)Scope: full doc + PR body, cross-checked against a completed primary-sourced research pass (Li & O'Donoghue 2014 taxonomy; Stephensen 2016 logit scaling; DYNASIM/MINT/POLISIM practices; determinism recipes; Trustees intermediates), against Verified sound
FindingsSHOULD-FIX 1 — the strongest method in the literature is never adjudicated (logit scaling, Stephensen 2016). §2.1 adjudicates only the weak member of the probability-adjustment family — annual linear probability-ratio scaling (the DYNASIM3 precedent) — and rejects it on three grounds: expected-not-realized margins, clipping at one, no unique displaced set. Logit scaling (IJM 9(3) 2016: multiply odds, not probabilities, by a common factor solved against the weighted margin) is exact on weighted expected margins, KL-minimal, deterministic in its factor, has a weighted variant, and cannot clip — it defeats the second objection entirely, and with a common-uniform reuse rule it also yields a unique deterministic flip set, weakening the third. The doc's grep-clean absence of "Stephensen"/"logit" means the selection-vs-scaling adjudication knocks down a strawman. The chosen mechanism plausibly still wins — logit scaling controls the expected weighted count, not the realized one, and realizing a count from scaled probabilities requires a selection step anyway — but that decisive ground must be argued against the strong method, not the weak one. Bound up with this: the registered flip ordering ( SHOULD-FIX 2 — the V.A2 claim "no unique gross addition/removal decomposition" is false against the pinned table. Verified on the live 2026 single-year table: projected years publish the full column set for both categories — LPR inflow / outflow / adjustments-of-status / net, and temporary-or-unlawfully-present inflow / outflow / adjustments-of-status / net (2030 intermediate: 600 / 263 / +450 / 788 and 970 / 329 / −450 / 191; total net 978, arithmetic exact), with footnote e making AoS a pure cross-category transfer. Gross additions and removals are uniquely decomposed at year × legal-status granularity; what the table genuinely lacks is age-sex cells, the Social-Security-area-to-roster bridge, and within-category flow semantics (fn c mixes citizen emigration into LPR outflow; fn b's residual method conflates temporary-stock outflow components). The error is conservative in direction (it defers a margin the table partially binds), but a design whose §3.3 discipline is "definitions must not drift" cannot misdescribe its own pinned source, and decision 2's premise ("implement V.A2 net change without an arbitrary split") overstates the gap — the split is published; the age-sex disaggregation and bridge are what is missing. Rewrite the §3.2 row-1 binding and the decision-2 text accordingly. SHOULD-FIX 3 — every m6-doc line anchor is stale against the merge target. The PR branch base (75d30dd) predates merged #216, and master's SHOULD-FIX 4 — seam ownership and the concrete gate-isolation mechanism are unnamed (a ninth open decision in substance). §4.1 says the wrapper "inserts reconciliation hooks at the named seams" of a loop that today has no seam surface ( NOTE 1 — publish the implied per-cell alignment factor as a first-class field. §5.2 publishes target, raw, and aligned totals plus raw residual per (margin, year, stratum), so the misspecification diagnostic is derivable — but the literature the doc itself cites states the diagnostic in factor terms (Toder's large-factor warning; DYNASIM's factors-as-diagnostic practice), and §6.3's corridors use shares and absolute displacements only. An explicit NOTE 2 — key tie values on the unit, not on frame position. §4.4 derives tie streams from (draw index, spec hash, year, margin, stratum, revision) but does not say the per-unit tie value is keyed on NOTE 3 — the emigration seam placement is decided without rationale while decision 2 is open. §3.2 fixes "a separately registered emigration branch runs after addition and before mortality" although the entire migration mechanism (units, donor pool, joint vector objective) is deferred to decision 2. Either state the ordering rationale (single-transition-per-year discreteness would do) or fold the placement into decision 2 so the future ratification is not half-pre-committed. VerdictREVISE. No blocking findings: the certification law is faithfully pinned, the existing accounting function is characterized code-exactly, the determinism recipes match the strongest published practice, all fifteen external citations verify against primary sources, and the doc's honesty about its own residuals (floors, non-exactness, unavailable margins) is exemplary. Revision is required because the central mechanism adjudication omits the strongest method in the literature (SHOULD-FIX 1), one pinned-source factual claim is wrong (SHOULD-FIX 2), the m6-doc anchors break on the merge target (SHOULD-FIX 3), and the gate-isolation enforcement needs a named structural mechanism (SHOULD-FIX 4). 0 BLOCKING / 4 SHOULD-FIX / 3 NOTE. |
Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Publish the implied per-cell alignment factor and key tie values on stable alignment units. Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Bind the published gross flow columns, preserve their unresolved roster bridge, and explain the emigration seam ordering. Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Record the two orchestration architectures and require a closed unaligned gate facade with structural and byte-identity enforcement. Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Update every mutable M6 document line pin to the merged post-3h master tree while retaining the immutable amendment blob link. Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Distinguish adjustment-of-status state changes from roster additions and removals throughout the migration design. Co-Authored-By: Codex gpt-5.6-sol <noreply@openai.com>
Fixes appliedApplied the alignment-design referee findings from comment 4986136518. Finding → commit map
Validation
No |
Summary
afterprojection around the existingbuild_alignment_displacementaccounting functionReview posture
This is a docs-only design for adversarial review. It changes no code, gate, floor, threshold, or run artifact. It does not resolve the separate re-drawn-
T*null-with-reason surface.The design incorporates independent contract and citation reviews, including exact-key accounting across roster-changing paths, homogeneous-unit calls to the existing function, seam-time threshold-prefix resolution floors, fresh replay state, chained earnings updates, stable synthetic IDs/ordinals/RNG coupling, and joint-vector deferral for atomic multi-cell migration units.
Checks
git diff --checkquarto pandoc docs/design/alignment_machinery.md --from gfm --to html