Revision log: 03-late-lessons-and-huang.md#
Revision of 26 September 2026 in response to three reviews in this folder: fidelity.md (F), balance.md (B) and completeness.md (C). Each issue was checked against the sources named in the review (transcript, report text extracts, LLA, HA and the working files) before it was acted on. Line numbers in the reviews refer to the 11:32 version (917 lines); the revised file is about 1,085 lines. No passage was shortened substantially; the overlap trims C suggested were not made (see C-length).
Stand-alone requirement: the revised document contains no article angles, addresses no particular reader, and replaces project-internal apparatus with named companion documents and two new appendices. The separate angles article the user asked for was not part of this revision task and has not been created here.
Fidelity review (F)#
| # | Issue | Verification | Outcome |
|---|---|---|---|
| F1 | M1 overstated as “at least as much harm as bad faith” (§4.9) | LLA §6.11 M1 Strength: relative size never measured; Limits: documented bad faith behind some largest harms | Fixed. Reworded to “at least as common a source of delay… relative size never measured… documented bad faith lies behind some of the largest harms” |
| F2 | “who bore which costs and benefits, and when?” printed as LL1 quotation (§4.6) | LL1-00 text, p. 11: “including their distribution between groups and across time” | Fixed. Actual wording quoted |
| F3 | “can make an incentive feel like sincere belief” printed as LL2-25 quotation (§8.1) | Phrase is LLA §4.3’s; LL2-25 p. 614 says people “engage in self-deception that helps them reinterpret or disguise” | Fixed. Report’s words quoted; paraphrase attributed to LLA §4.3 |
| F4 | MTBE working-note phrases attributed to “the reports’ notes” (§4.11) | Phrases in notes/LL1-11.md only; LL1-11 text has “a technical problem that can be managed by risk reduction” and “the possibility of risk reduction being insufficient” |
Fixed. Report text quoted (pp. 115, 117); the containment contrast given as the companion analysis’s reading, unquoted |
| F5 | Narayanan and Kapoor ellipsis drops “and the risk of brand damage”; “about liability alone” wrong | E4 L113 confirms full sentence | Fixed in §2, §4.3, §7.1 and §11.2; “close to” changed to “overlaps with” their proposals |
| F6 | “research dialogue” given an interview timestamp (§10.4) | Not in transcript; Dwarkesh Patel, April 2026 (HA §4.8, E1) | Fixed. Attributed; interview words quoted with [1:37:36] |
| F7 | Cohort statistic described as “19% below trend” (§7.2) | D06 L172: relative gap to less-exposed peers, descriptive | Fixed |
| F8 | Vinyl chloride figures used uncorrected (§4.6) | Hindsight LL2-08: like-for-like overestimate about fourfold | Fixed in §4.6 and Appendix A |
| F9 | OpenAI 26 June compromise stated as fact (§4.8) | HA Appendix B: sequence disputed across files | Fixed. Point rests on Hugging Face’s detection and the Australian breach; dating noted as disputed |
| F10 | Beryllium “lesson” misattributed (§11.3) | LL2-06 p. 140 (“must be discounted”) v. Guidotti p. 148; hindsight: producer co-drafted tighter limit | Fixed. Main authors’ view and dissent both stated; hindsight partly vindicates dissent |
| F11a | Fisheries triggers “under industry pressure” (§10.2) | Hindsight LL2-17: “sometimes scientifically justified… also a channel for pressure” | Fixed |
| F11b | Cod “irreversible demise… overturned” (§6.1 item 11) | Hindsight LL1-02: “Healthy” status rests heavily on revised LRP | Fixed. K5 caveat added |
| F12 | S7’s two case families called “both natural hazards” (§4.8) | LLA S7: nuclear accidents and floods | Fixed |
| F13 | “influenced” printed as interview speech / unsourced (§4.4, §7.1, §11.2) | Not in transcript; All-In, 14 September (E1 L251) | Fixed at all three places; only “influenced” in quotation marks |
| F14 | “in silence” and “two out of three rights” unsourced | HA L602 (All-In); HA L478, E1 L137 (Lex Fridman, March 2026) | Fixed. Sources added (first use in §4.1 and in §6.1) |
| F15 | Anthropic’s 6–12% read as OpenAI’s share (§4.3) | HA L866: Anthropic’s own compute | Fixed |
| F16 | Klein “overstated” rates attributed to FC C097, which rates him mostly accurate (§4.1) | FC C097 verdict; HA S3 notes on the misplaced quotation mark | Fixed. Mirror now uses the Apollo misattribution (C097) and flags the gloss’s attribution as probable Klein paraphrase |
| F17 | “These products weren’t released” relied on without attribution flag (§4.8) | HA §1.4: “almost certainly Klein’s“ | Fixed. Flag added |
| F18 | Leaded-petrol clearance dated 1925 and 1926 inconsistently | LL2-03 text: committee “declared that TEL was safe for general use, in 1926” | Fixed. 1926 used for the clearance throughout (Hamilton’s 1925 conference remark kept) |
| F19 | Ethyl, Monsanto and beryllium producer cast as Huang’s precedent (§3.2, §4.11) | Contradicts §8.1; LL2-27 p. 647 supports a value-chain finding without naming concealing firms | Fixed. Named firms replaced with the value-chain finding and a position-not-conduct caveat (see also B7) |
| F20 | M6 missing from LL2-22 disclosure list (§1.5) | LLA M6 evidence cites LL2-22 p. 543 | Fixed. M6 added; LL2-22 entries marked † in Appendix C |
| F21 | Smaller points: “reports’ rule 5”; G2 among [K] entries; “two labs”; LL2-21 p. 520 for expert-witness roles; “under inquiry”; OpenAI pre-emption unconditional; Box 20.4 applied to pre-emption; hindsight quote cited to p. 80; OpenAI competitor clause conditions; 13–15% reconstruction; recording date; “six families” | Each checked (E3 10-Q wording; HA §2.3; LL2-20 Box 20.4 text; D12 L189; FC C002; HA §1.4; LLA §5.6; D12 on Google’s May incident) | All fixed. Google’s incident added as a third lab with its confirmation date and “reported” |
| F22 | Stand-alone use: folder path in Appendix B; unpublished apparatus | Confirmed | Fixed. Companions named by title (and filename in Appendix B); folder path removed; hindsight and FC conventions explained; Appendices C (lens key) and D (fact-check verdicts) added |
Balance review (B)#
| # | Issue | Verification | Outcome |
|---|---|---|---|
| B1 | “Does not see” evaluation awareness and pre-release harm contradicts his own statements | Transcript [36:44], [44:17], [48:58], [53:36] | Fixed. §4.8 “without reconciling” removed and restated; §4.9 frame paragraph rewritten (“sees… where it stops”); §5.3 K2 row recast around the disciplining mechanism; §8.3 H3 and H6a reworded; §8.4 qualified to “implication for release decisions” |
| B2 | “I know they know how to fix it” not split by sub-question; BSE parallel conflates two questions | Transcript [44:17]→[55:46]; FC C097; Anthropic 9 September findings | Fixed. Split defined once in §3.4 and applied in §2, §4.2, §4.9, §5.3 (T1 row), §7.2; BSE parallel restated on the behavioural limb and lowered to medium-low |
| B3 | Moral-hazard argument missing; [47:10] and “safety is paramount” absent; Altman UN sentence absent | HA §7.4 item 2; §7.3(b), (d); transcript [47:10], [44:17] | Fixed. Moral hazard added to §2, §4.7, §6.1 (item 16), §6 analysis; §8.2 pattern 6, §8.4, §9.5, §4.9 reworded; Dreamforce quotation paired with [47:10] in §3.4 and §4.7; Altman added to §6.1 item 10. “Safety is paramount” not added separately (its [44:17] turn is already quoted for the containment diagnosis) |
| B4 | In brief implies a field-wide gap is Huang’s own, and says he lacks an outside check he endorsed | Transcript [51:20]; Astra card | Fixed in §2, §4.1 and §7.1 |
| B5 | “0% chance” compared with estimates of a different event and horizon | HA T8 caveat; FC C124 | Fixed in §2, §3.4, §4.2 (finding and Mirror), §5.3, §6.1 item 2, §9.1; “near zero, not 0%” kept in §11.2 as a point about form |
| B6 | Shutdown “liabilities” read as costs on the declarer without noting the more natural reading | Transcript [36:44] | Fixed. Full sentence quoted in §4.2, both readings stated, declarer-pays point rested on the cost of shutdown itself, confidence kept moderate |
| B7 | Historical analogies cast Huang as concealing manufacturers; caveats dropped | §8.1, §4.5 of 03; LL2-27 p. 647 | Fixed. §3.2 and §4.11 recast; “structure, not conduct; one case, moderate weight” attached to DuPont (§2, §4.7, §7.1) and Kettering–Midgley (§4.9); DuPont outcome clause marked as outcome; pet-food counter-case added to §7.1 |
| B8 | Dependence on unpublished apparatus and undefined identifiers | Confirmed | Fixed. Rules numbered 0–10 in §1.3; companions named; Appendices C and D; FC codes explained. Primary-source citation for Montzka already present |
| B9 | Ranks 3 and 4 double-count the firm-held gate; I5 stretched to Huang; EO title as evidence | LA3 I5 record; I5 strength rests on public bodies | Fixed. Firm-held gate merged into rank 3; rank 4 limited to the promoting state with Huang’s role stated as advisory; In brief amended; EO title no longer used as evidence; “completely aligned” cut to one use |
| B10 | Export-control error allocation one-sided; loosening read as C1 confirmation | Transcript [1:35:15] | Fixed. Both error costs stated as each side frames them; loosening “consistent with C1, not evidence of capture” |
| B11 | 72-entry table adds up and drops direction; no baseline | LA2 and LA4 direction columns; other LA records’ “In his favour” | Fixed. Total row deleted; paragraph added listing entries that cut for Huang or both ways, and stating no critic baseline exists |
| B12 | Mirror claimed two-sided but not shown with comparable structure; some Mirror cells not mirrors | Confirmed | Fixed. Compact table “where Late Lessons presses on the critics” added to §5.5; In brief softened; rank 5 and rank 8 Mirror cells replaced with true mirrors |
| B13 | Narayanan and Kapoor’s “We were wrong” used as a verdict | E4 §2.3 | Fixed. What they revised, their continued security reading and their non-pacing remedies stated in §2, §4.3, §7.1 |
| B14 | §8 discounts Huang’s public words more than others’ | Transcript [15:04] in context | Fixed. Comparative removed; applied to all participants |
| B15 | Timing parenthetical and “while” clause invite motive-from-outcome inference | §8.1 rule | Fixed. Parenthetical tied to L4; §4.4 sentence split with the rule-0 caveat placed first |
| B16 | “In silence” unsourced and out of context | HA T13; E1 | Fixed. Attributed, 2025 statement added, reading marked contested, confidence lowered to low |
| B17 | Over-corrections and inconsistencies (Coxon; early dates; shutdown condition; interest symmetry) | E4 L246; §8.5 of 03; transcript [36:44] | Fixed. Coxon’s first response added (second-hand); §8.5 formulation used in §4.4 and §9.2; shutdown condition weighted as weak evidence in §4.4, §8.3, §9.2; “no more closely” replaced |
| B18 | Klein Mirror misattributed; cites C097 against its rating | Transcript [48:21]; FC C097, C100; HA S3 | Fixed (with F16) |
| B19 | Weak items carried at high confidence in K1 | FC C098 “Accurate”; “did no harm” context unknown | Fixed. C098 noted; “did no harm” reading lowered to medium |
| B20 | Word-level inferences (“hurt”, “risk”, “who decides”, “master move”) too strong | Transcript [55:46], [1:37:36], [15:04], [1:19:12] | Fixed. Pattern 3 renamed with the split; “risk” count presented as colour; pattern 4 narrowed; “most characteristic move” |
| B21 | Workflow mechanics show through; production method not stated | Confirmed | Fixed. §1.2 rewritten as method, including a statement that the analysis was prepared with extensive AI assistance; “first-pass”, “integration”, “was run” removed |
Completeness review (C)#
| # | Issue | Verification | Outcome |
|---|---|---|---|
| C1 | No key to 70 cited entries; rules unnumbered; internal references | Confirmed | Fixed (with B8, F22) |
| C2 | Reports’ support for provisional measures missing; evaluation awareness’s effect on pause exits missing | LLA §6.12; D01 §6; D03 L158; LL2-12 p. 274 (wording checked: “can easily become an essentially scientific or academic pursuit”) | Fixed. Added to §2, §4.1 Mirror, §5.5, §11.1 (new row), §11.2, §12.1 Q1 |
| C3 | Who decides, and the public’s place (I10) missing | LA3 I10; LL2-28 p. 671; transcript [40:21], [51:20], [1:27:32] | Fixed. New finding in §4.7; lines in §7.2, §12.2, Appendix A |
| C4 | Recursive self-improvement and the relocated human in the loop thin | D08 §4.4; HA T11; transcript [1:12:47], [1:15:30], [1:15:35] | Fixed. New finding in §4.8; §11.2 bullet; §12.1 question 9 |
| C5 | Open weights against the release gate scattered | HA T12; D12 (Nemotron framework “not in the record”, so worded that way rather than “none found”) | Fixed. Consolidated finding in §4.8; §11.2 bullet; §12.1 question 10 |
| C6 | Points in Huang’s and critics’ favour dropped (0% caveat; Hinton’s survey range; attributability; I7 alignment; DeepSeek record; Tabarrok) | HA T8, FC C124; D10 §6 item 8; LA3 I7; D08 L264; leaders comparison | Fixed. All six restored (§3.4, §4.2, §4.3, §4.4, §4.5, §6.1 items 2, 5, 17, 18, §9.1). The review’s GM “forty years” limit on I7 was not used (not verified); a different limit (I3) given instead |
| C7 | §11.1 and §11.3 shorter than the evidence supports | D03, D06, D07, D08, D09, D10 §7 lists | Fixed. Four rows added to §11.1; nine rejectables added to §11.3, burden-of-proof item flagged for LL2-22 |
| C8 | Prior justification of uses absent | LLA §6.12; LL1-03, LL1-16 p. 176; FC C010 | Fixed. §4.7 support, §11.1 row, §11.2 “On uses” |
| C9 | Stakes of outside evaluators and commentators not examined | Rule 0 | Fixed. Caveat in §1.6; “independent” changed to “outside” where independence was unchecked (§2, §4.8, §4.12, §6.1, §10.3). METR’s investigation kept as “independent”, its standard description in the sources |
| C10 | Wider landscape almost entirely American | Confirmed | Fixed. Caveat in §1.6, bullet in §10.6, question in §12.2. No new research, per the task’s rule 8 |
| C11 | Financial and systemic-economic layer fragmentary | Transcript [1:21:05] (“collateralized”), [1:29:48]; FC C176; D05, D06 | Fixed. “Financial coupling” finding in §4.5; Ratepayer Protection Pledge and Microsoft in §4.6 |
| C12 | History, learning and skills exchanges under-used | Transcript [13:44], [15:04], [21:16], [22:26], [24:52] | Fixed. Second non-engagement instance in §8.4 and §8.3; K10 learning finding in §4.6 (low confidence); §12.1 question 11 |
| C13 | Water and shared resources missing | FC C209; D08 S2/S6 | Fixed. §4.6 water sentence; S6 evaluation-validity point in §4.8; §6.1 item 14 qualified “on energy” |
| C14 | Per-leader M1 insulation and position shifts missing | Leaders comparison §M1 and table 2D | Fixed. §9.4 M1 paragraph; §9.2 pattern 5 on shifts |
| C15 | K6 and K3 unused | LA1 K3, K6 records | Fixed. §4.1 finding and Mirror; K6 row in §5.3; K6 cited in §8.5 Mirror |
| C16 | Smaller omissions | Transcript [1:34:16]; D08 §4.16; D07; D10 | Fixed: security question (§4.10), cross-layer interactions (§4.8), incident-trend rationale (§11.2), US–China incident mechanism (§10.4, post-recording, tentative). Not adopted: the Great Lakes point and the Sega candour story, both marked optional by the reviewer; they add little to findings already made and would lengthen the text without changing a conclusion |
| C-length | Trim overlap (4.12 v. 10.1–10.3; Appendix A; repeated “no exits” Mirrors) to offset additions | Task instruction: do not shorten substantially | Not adopted. Sections are kept as stand-alone entry points for readers of a general resource; the document grew by about a quarter (much of it the two new appendices) |
Unresolved or for the author’s judgement#
- AI-assistance statement (§1.2). Added on B21’s recommendation as a factual description of how the document was produced. Its wording is for the author to confirm.
- Companion filenames. Appendix B gives the current filenames of the two companion analyses; these should be updated if the files are released under other names.
- Unpublished working analyses. Appendix B describes them as “available on request”. If they will not be made available, that phrase should be removed.
- Separate angles article. Not created in this revision; the document itself contains no angles.