Gavin Baker: The weekend event map shows Dario as most maximalist, while Elon peer-review, Demis FINRA, and Sacks unilateral pace are wildly different packages
The Gist
Baker says do not mash the weekend into one story. Dario asked for the biggest stack of rules (though less than his old FAA idea). Elon meant rival labs checking each other. Demis wants a FINRA-like body. Sacks says just slow yourselves down and drop the cartel talk. Others are arguing about who counts as a real independent evaluator. Steelmanned reconstruction of Gavin Baker's Sep 13 2026 X note for LogicFirst analysis; not an endorsement of Atreides Management views or investment advice.
Conclusion
The weekend's proposals are not one consensus: Dario's package is the most maximalist (though milder than his prior FAA-for-AI), while Elon's competitor peer-review, Demis's FINRA-like SRO, Sacks's unilateral pace, Sriram's independence demand, Clem's HF evaluator offer, and Wang's alignment focus are distinct and often conflicting packages.
Premises
- Dario's weekend package is the most maximalist among named proposals: embedded evaluators, a national regulatory regime above capability/ingredient thresholds, a democratic international pact, stricter China compute/distillation limits, a separate global regime that includes China, and a Sherman Act waiver so frontier labs can coordinate before a national regime exists.
- That package is still less maximalist than some of Dario's prior asks (e.g., an FAA-for-AI in Policy on the AI Exponential); Baker treats him as sincere and also notes the package would likely be good for Anthropic's long-term business.
- Sam's concrete match is the evaluator step (liability-aligned), not endorsement of every maximalist limb.
- Elon's Dario is right clarified into competitor peer review (MPAA-like regular calls plus 1-2 week pre-release safety review by rivals), which is wildly different from embedded independent evaluators plus national/global regimes, and aligns more with David Sacks; Elon also said open-weight models will not be slowed.
- Demis called Dario's essay a step in the right direction and has pushed a FINRA-like SRO; Dario said he is open to that as part of his proposal.
- Sacks urged unilateral pacing, called the antitrust waiver a cartel request, and denied METR independence given Anthropic investor/staff intertwining; Sriram Krishnan stressed unaffiliated independent evaluators (an indirect METR dig); Clem/Hugging Face offered to serve as a neutral third-party evaluator; Alexandr Wang said alignment will be an increasing focus.
- Accurate policy and market reading therefore requires disaggregating these packages rather than treating weekend AI safety consensus as one blob.
Assumptions
- Maximalist is comparative within the weekend set and relative to Dario's own prior FAA ask, not a claim that Dario seeks a total AI ban.
- Sincerity and business-alignment can coexist.
- Baker's if Jensen was consulted aside about Clem is speculative color, not load-bearing.
- Research residual: METR states it takes no frontier-lab cash and has publicly criticized Anthropic analyses; Sacks/Sriram still treat investor/alumni intertwining as independence failure. Elon's July Economist peer-review sketch matches Baker's MPAA reconstruction.
- Differs: METR's no-lab-cash policy and critical Anthropic reviews qualify a cash-capture reading; Baker's steelman retains the independence dispute as politically load-bearing.
Analysis
Overall strength: Moderate. Argument type: Inductive.
Premise Strength
- Dario's weekend package is the most maximalist among named proposals... (Moderate) — Individual policy elements are well-documented and specific, but the superlative ranking depends on an implicit, unweighted comparison metric; the claim's scope is helpfully bounded by A1 (comparative, not absolute) which mitigates but doesn't eliminate the arbitrariness risk.
- That package is still less maximalist than some of Dario's prior asks...Baker treats him as sincere and notes business alignment (Moderate) — The comparison to the prior FAA-for-AI ask is verifiable and reasonably clear; the sincerity attribution, however, is an unfalsifiable psychological claim, and A2's move to declare sincerity and self-interest compatible resolves a false dichotomy without independently establishing sincerity.
- Sam's concrete match is the evaluator step, not endorsement of every maximalist limb (Moderate) — A useful and plausible narrowing that blocks an illicit inference from partial to full endorsement, but it functions as Baker's interpretive gloss on Altman's position rather than a directly quoted disclaimer.
- Elon's position clarified into competitor peer review, wildly different from embedded evaluators/national regimes; aligns with Sacks; open-weight unslowed (Strong) — This is the argument's best-evidenced contrast: a concrete, named mechanism (MPAA-style peer review) is set against a concretely different mechanism (embedded evaluators), and is independently corroborated by a prior Economist interview per A4.
- Demis called Dario's essay a step in the right direction and pushed a FINRA-like SRO; Dario open to it (Weak) — This premise actually demonstrates partial convergence rather than conflict, sitting in tension with the overarching 'wildly different packages' framing applied to the full set.
- Sacks/Sriram/Clem/Wang positions (unilateral pacing, cartel framing, independence critique, neutral evaluator offer, alignment focus) (Strong) — These are explicit, mutually contradictory characterizations (cartel vs. coordination; captured vs. independent evaluator) directly sourced to each figure's own statements, making this the strongest evidentiary cluster for genuine disagreement.
- Accurate policy and market reading therefore requires disaggregating these packages (Moderate) — A reasonable inductive summary given the preceding contrasts, though it is a derivative claim whose strength is capped by the weaker links above (P1's ranking metric, P5's convergence, single-source aggregation) and overstates uniformity of 'wildly different' across all pairs.
Potential Fallacies
- Unweighted composite ranking (P1 / Conclusion) — The claim that Dario's package is 'most maximalist' aggregates several disparate policy dimensions (institutional reach, compute restrictions, antitrust waiver, international scope) without a stated weighting scheme. Reasonable people using different weights (e.g., prioritizing impact on open-weight models) could reorder the ranking, meaning the central comparative claim is more of a qualitative judgment call than an objective measurement.
- Asymmetric charity (selective motivated-reasoning scrutiny) (P2/A2 vs. P4-P6) — Dario's proposal is explicitly granted 'sincere despite business benefit' status, but the same self-interest lens is not extended to Sacks's anti-waiver stance, Elon's open-weight carve-out, or Wang's alignment framing, even though each plausibly serves the speaker's competitive position. This uneven application of charity is not a formal error but weakens the argument's claim to neutral disaggregation.
- Unresolved rebuttal treated as still load-bearing (P6 / A4-A5) — The argument acknowledges that METR takes no frontier-lab funding and has criticized Anthropic (a direct rebuttal to the 'captured evaluator' claim), yet stipulates that the independence dispute remains politically load-bearing without explaining why the rebuttal is insufficient. This surfaces a live defeater without resolving it on the merits.
- Non-independent evidence aggregation (P1-P6 collectively supporting P7) — All six comparative premises are drawn from and interpreted by a single commentator synthesizing a single news cycle. Treating them as multiple converging data points risks overstating the diagnosticity of what is, in evidentiary terms, one analyst's reading of public statements.
Counterarguments
- P1 / Conclusion (High impact) — Reweighting the maximalism criteria (e.g., prioritizing restrictions on open-weight distribution over number of institutional asks) could just as easily cast Sacks's deregulatory unilateralism as 'maximalist' in the opposite direction, showing the ranking is an artifact of chosen criteria rather than an objective fact about the proposals.
- Conclusion / P7 (High impact) — All six figures arguably converge on a shared underlying principle - some form of pre-deployment evaluation before frontier release - with disagreement concentrated on institutional mechanism (who evaluates, under what authority) rather than on whether evaluation should happen at all. Framing this as 'wildly different packages' rather than 'shared principle, differing implementation' overstates fragmentation.
- P2 / A2 (Medium impact) — Rather than sincerity and business-interest coexisting neutrally, Dario's package could be read as safety-washing: an expansive regulatory ask (national regime, international pact, compute limits) that would raise compliance costs disproportionately for smaller/open competitors while entrenching Anthropic's first-mover advantage.
- P6 / A4-A5 (Medium impact) — METR's stated no-lab-cash funding policy and its documented public criticism of Anthropic directly rebut the claim that its independence is compromised by investor/staff ties; treating the dispute as still 'load-bearing' after conceding this evidence understates how much the rebuttal weakens Sacks/Sriram's position.
- Overall framing (Medium impact) — Publicly emphasizing fragmentation ('no consensus') can itself function as a strategic move that industry actors and their allies use to forestall coordinated regulatory action ('even the labs can't agree, so nothing should be mandated'), meaning the disaggregation project, however analytically accurate, may have a real-world effect of delaying oversight regardless of intent.
Suggested Improvements
- Maximalism ranking methodology — Specify an explicit rubric (e.g., scope of authority, enforcement mechanism, geographic reach, reversibility) and weight for scoring each proposal's stringency before declaring one 'most maximalist.' This would convert a qualitative, potentially cherry-picked judgment into a transparent, falsifiable comparison that critics could contest on the merits rather than dismiss as arbitrary.
- Symmetric self-interest scrutiny — Apply the same 'sincere but also good for [X]'s business' lens evenhandedly to Sacks (a16z/administration ties), Elon (xAI/open-weight positioning), and Wang (Scale AI/Meta positioning), not only to Dario. Symmetric treatment would strengthen the argument's claim to neutral disaggregation and preempt charges of selective charity.
- Source verification — Supply direct quotes or links for each attributed position (Altman's evaluator-only endorsement, Elon's peer-review clarification, Demis's SRO openness) rather than relying on the author's paraphrase. Primary-source citation would allow independent verification and reduce dependence on a single interpreter's fidelity in summarizing seven people's positions.
- Principle vs. mechanism distinction — Explicitly separate 'do all parties agree evaluation/oversight of some kind is needed' from 'do they agree on the institutional form it should take,' and frame the conclusion around the latter. This would preserve the paper's valuable disaggregation of mechanisms while avoiding the misleading impression that there is no shared ground at all, which is vulnerable to a strong steelman rebuttal.
- Resolving the METR dispute — Either adjudicate the independence question using METR's funding disclosures and staff/alumni network data, or explicitly flag it as an empirically resolvable question requiring further disclosure rather than treating it as a permanent political fault line. Leaving a defeater acknowledged but unresolved weakens the argument's claim to being an evenhanded, fact-based steelman of the independence controversy.
Scenario Tests
- Maximalism criteria are reweighted to prioritize restrictions on open-weight model distribution rather than number of institutional asks (Challenges) — Dario's ranking as 'most maximalist' could be overtaken by Sacks's or Elon's positions, showing the central comparative claim is sensitive to unstated methodological choices rather than being an objective fact.
- All six proposals are recoded around the shared premise that some pre-deployment third-party evaluation is desirable (Challenges) — The 'wildly different packages, not one consensus' framing weakens into 'shared principle, differing institutional mechanism,' which is a real and defensible alternative reading of the same facts.
- Independent audits confirm METR's no-lab-cash policy is robust and Sacks/Sriram's independence critique is largely rhetorical (Challenges) — The 'politically load-bearing' framing of the independence dispute in A4/A5 would collapse, exposing an asymmetry in how evenhandedly the argument treats competing narratives.
- Dario's current package is compared strictly against his own prior FAA-for-AI ask, isolated from cross-actor comparison (Supports) — The 'milder than his prior ask' claim (P2) holds up well as an intra-personal, verifiable comparison, which is one of the argument's more solid empirical anchors.
- The weekend's statements are revisited a few weeks later after further negotiation/clarification (Neutral) — Given the ephemeral, social-media-sourced nature of the evidence, positions may shift quickly, meaning the 'map' is a time-bound snapshot whose specific claims (though not necessarily its general disaggregation thesis) could become stale.
Coherence & Relevance
The argument is internally coherent and methodically organized as a comparative taxonomy: each named figure's position is given a discrete, sourced characterization, and the conclusion follows as a modest inductive generalization rather than a claim requiring strict logical entailment. Its main coherence gap is the tension between the Demis/Dario premise (P5), which shows partial convergence, and the blanket 'wildly different packages' framing applied to the whole set in the conclusion. The argument's reliance on a single interpreter's synthesis of ephemeral public statements, combined with an unweighted 'maximalist' ranking and asymmetric application of the sincerity-business-alignment defense, means the overall disaggregation thesis is well-supported in its broad strokes (genuine mechanism-level disagreement exists) but should be held with moderate rather than high confidence in its specific comparative rankings and its claim to neutral, comprehensive coverage of the underlying policy space.
- Dario's package details (P1) (Strong) — Directly establishes one pole of the comparative spectrum, but the 'most maximalist' superlative requires an unstated weighting method to be fully justified.
- Comparison to prior FAA-for-AI ask and sincerity/business note (P2) (Moderate) — The historical comparison is solid; the sincerity claim is an interpretive addition not strictly derivable from the comparison and is applied asymmetrically relative to other actors.
- Sam's narrow evaluator-only match (P3) (Moderate) — Supports the disaggregation thesis but rests on Baker's inference about Altman's intent rather than a quoted disclaimer.
- Elon's peer-review clarification (P4) (Strong) — One of the most concretely evidenced contrasts in the argument; minimal gap given corroboration via the Economist interview (A4).
- Demis's SRO push and openness (P5) (Weak) — This premise reads more as evidence of convergence than conflict, creating an internal tension with the 'wildly different' framing applied to the conclusion.
- Sacks/Sriram/Clem/Wang positions (P6) (Strong) — Provides the clearest evidence of substantive disagreement; the METR independence sub-dispute (A4/A5) is acknowledged but left unresolved.
- Disaggregation conclusion (P7) (Strong) — Follows reasonably from the cumulative contrasts but inherits the weaknesses of its weakest supporting premises (P1's ranking metric, P5's convergence) and slightly overstates uniformity of divergence across all six actors.