Cognition SWE-2 - Near-Frontier Coding Model for Devin Agents

Cognition SWE-2 is a coding model for engineering teams that want frontier-level software-engineering performance without paying frontier token prices. Released on September 10, 2026 inside Devin Desktop and CLI, SWE-2 is post-trained from Moonshot AI's Kimi K3 and routed through the Devin product rather than sold as a standalone API.

Core Features

  • Reinforcement-learning post-training from Kimi K3 (a 2.8-trillion-parameter base) tuned specifically for code generation and task cost.
  • Three effort levels - medium, high, and max - trained in a single run with a cost penalty, so you trade compute for harder tasks without switching models.
  • Near-frontier scores: 50.0% on FrontierCode 1.1 Main (within one point of Anthropic's Claude Fable 5.1) and 92.8% on Terminal-Bench 2.1.
  • Fewer steps to value: a median of 18 steps to the first real edit, down from 48 in the earlier SWE-1.7.
  • Bundled access: SWE-2 runs inside Devin Pro, Max, and Teams with no separate token rate.

Use Cases

  • Routine PR reviews, small features, and bug fixes where cost per task matters more than marathon agentic runs.
  • High-volume coding shops already on the Devin platform that want cheaper model calls.
  • Teams comparing completed-task cost rather than raw benchmark leaderboard position.

Pricing

SWE-2 has no standalone API price; it ships inside Devin. Devin Pro costs $20 per month, and Cognition offered SWE-2 free across Pro, Max, and Teams subscriptions through October 10, 2026. See related models in our AI engine and model directory.

Our Take

Best for budget-conscious teams doing short, well-scoped coding tasks inside Devin. The trade-off is clear on long-horizon work: on Terminal-Bench 4 it scores 27.3% versus 55.8% for Claude Fable 5.1, so overnight multi-file agentic sessions still belong to a frontier model.

FacebookXWhatsAppEmail