Story: state of ai/safety
Context: SB 1047’s author, responding to the veto — a checkable claim about the reliability of voluntary commitments. Partially borne out by the record on this page (OpenAI’s dissolved Superalignment team and unfulfilled compute pledge), but ‘rarely work out well’ is a general assessment not settled by any single instance. Moved from unverified to mixed this cycle. The direction of the claim is now documented: the Future of Life Institute’s Summer 2026 AI Safety Index finds Anthropic, OpenAI, Google DeepMind and Meta have all weakened or voided their unilateral-pause pledges, some substituting competitor-contingent conditions, and separately that companies which had banned military applications have reversed course. That is an advocacy organisation’s expert-panel assessment rather than an independent measurement, and the claim’s quantifier — “rarely work out well” — is not operationalised and cannot be scored true or false as stated. Mixed: the pattern it asserts is now on the record, the strength of the assertion is not testable.