Story: state of ai/leading models
Context: The checkable core holds. The published weights are for a 2.8-trillion-parameter model, larger than any previously released open-weight model — corroborated independently by Tom's Hardware and, before the release, by Rahul Shome of the Australian National University in Nature, who said K3 "would be the largest open-weight model". Two caveats that do not falsify it: "3T-class" is a generous rounding of 2.8 trillion, and only 104B of those parameters activate per token. The trailing capability claim is separately corroborated — Artificial Analysis scores K3 at 57 on its Intelligence Index v4.1, four points off the highest score on that index — but the benchmark table Moonshot published with the weights is not: on it, K3 is run with Moonshot's own Kimi Code harness while rivals use Claude Code or Codex, and Moonshot discloses that Claude Fable 5 hit fallbacks on 35% of its SWE-Marathon tasks.