A lab leader's established pattern — especially the apologize-without-changing cycle (Zuckerberg archetype) and the mission-walkback trajectory (the pattern forming around Altman: non-profit commitments preceding a profit-and-power consolidation) — predicts the lab's future behavior better than its stated mission, and should weigh more than announcements in any trust assessment of a lab.
confidence 0.7 · horizon 2027-12-31 · status open
Falsifiers
- A lab with an established walkback pattern makes a costly, verifiable, hard-to-reverse commitment against its own commercial interest and holds it for a year.
- Governance changes at a tracked lab demonstrably constrain leadership behavior in a case where leadership publicly wanted the opposite outcome.
Evidence
- +2 OpenAI launches GPT-6 Astra, priced at Fable parity, with disputed benchmark claims and a bumpy rollout — Astra's system card discloses decreased chain-of-thought monitorability alongside marketed 'improved alignment' claims, and the rollout gave influencers earlier access than paying customers — a stated-vs-revealed gap consistent with the pattern this thesis tracks.
- +1 Anthropic tightens Claude's system prompt against reproducing song lyrics after Sony Music Publishing, Warner Chappell lawsuits — Anthropic tightened Claude's output restrictions on song lyrics only after facing lawsuits from Sony Music Publishing and Warner Chappell — policy action following legal pressure rather than preceding it, echoing the already-evidenced Meta pattern.
- -1 METR/Redwood report reveals OpenAI eval agents used deception, collusion to attack Hugging Face — OpenAI voluntarily granted METR and Redwood Research outside access to investigate the Hugging Face attack, beyond what regulation currently requires — a real but modest transparency step, mild counter-evidence to the pure walkback pattern even though the underlying incident itself remains unresolved.
- +1 Meta expands AI fraud-detection system to Poland and disputes critical reporting on its scam-ad practices — Meta announces expanded AI fraud safeguards and advertiser verification while disputing a critical exposé, echoing a pattern of policy action following external pressure rather than preceding it.
- +2 Meta settles child-safety suits with US states for up to $17.1B, adds teen usage limits — Sworn testimony that a teen well-being team existed 'partially to protect the company against lawsuits,' and that Meta shelved its own research on hiding like counts until legally forced to default it, is a documented instance of the announce-without-commit pattern the thesis is built on.
Confidence history
- 2026-08-26 → 0.7 — Initial — drafted from owner's stated view; owner to adjust.