Cognition SWE-2: 92.8% on Terminal-Bench 2.1
RL-post-trained from Moonshot's Kimi K3 base (2.8T MoE, 104B active, 1M context), SWE-2 is a closed coding-agent model with the top published score on the provider-run Terminal-Bench 2.1 snapshot. Notably, the headline number surfaced first through third-party eval trackers rather than Cognition's own launch page — check the harness and version before repeating it.