Skip to this page
THE SUPER INTELLIGENCE TIMESBY OPCELERATE NEURAL RSS
← Back to The Super Intelligence Times
Models Desk / SWE-2 / Source-backed briefing / 2026-09-12
← Back to The Super Intelligence Times
The Super Intelligence Times
Source Notes Desk
Cream stationery folder with gold cube and pen on walnut desk for SWE-2 briefing
Models Desk / Cognition / Posted Sep 10 / Briefing Sep 12

SWE-2 Claims 50% FrontierCode Near Fable. Devin-Only. No API.

Cognition’s Sep 10 post introduces SWE-2 as its most advanced coding model yet. Vendor claim: 50.0% on FrontierCode 1.1 Main, within one point of Fable 5.1, while being 64% cheaper. It runs inside Devin—not as a standalone API.

Quick answerCognition posted SWE-2 on September 10, 2026: vendor 50.0% on FrontierCode 1.1 Main (within one point of Fable 5.1, 64% cheaper), post-trained from Kimi K3 (2.8T). Available starting that day in Devin Desktop and CLI; rolling out on Devin Web and Fusion. No standalone API, no open weights, no per-token rate card on the primary. Foreign Devin SaaS ≠ private Alberta path.

On September 10, 2026, the Cognition Team introduced SWE-2—framed as its most advanced coding model yet and a push on the capability–cost frontier. The post’s headline vendor claim is 50.0% on FrontierCode 1.1 Main, within one point of Fable 5.1 while being 64% cheaper. Those scores are Cognition’s published vendor results, not an independent Alberta desk audit. Treat them as publicly reported claims and read the primary before you rewrite a coding-agent budget.

AI automation AlbertaAI consulting AlbertaAI training Edmontonprivate AI security
FrontierCode 1.1Vendor: SWE-2 50.0% — within one point of Fable 5.1, 64% cheaper.
BasePost-trained from Kimi K3, a 2.8T-parameter model.
Efficiency (vendor)SWE-2 medium vs SWE-1.7: higher score, 58% fewer turns, 81% less avg cost.
AvailabilityDevin Desktop + CLI today; rolling out on Devin Web and Fusion. No standalone API.

What Cognition listed for SWE-2

The primary says SWE-2 scaled RL into the multi-trillion-parameter regime on top of the SWE-1.7 training recipe, with an RL approach that trains all reasoning-effort levels in a single run. SWE-2 is post-trained from Kimi K3 (2.8T). Cognition also publishes a vendor comparison table covering FrontierCode 1.1 Main, DeepSWE 1.1, Terminal-Bench 2.1, and Terminal-Bench 4—use those lines carefully as vendor claims, not as independent benches.

Vendor benchmark table from Cognition’s SWE-2 post (not independent benches).
BenchmarkSWE-2Fable 5.1SWE-1.7
FrontierCode 1.1 Main50.0%50.9%42.0%
DeepSWE 1.173.0%67.4%37.7%
Terminal-Bench 2.192.8%91.4%81.5%
Terminal-Bench 427.3%55.8%7.6%

On behavior, Cognition reports that on FrontierCode 1.1 Main, SWE-2 medium scores higher than SWE-1.7 while taking 58% fewer turns and costing 81% less on average. First real edit median is 18 steps for SWE-2 medium versus 48 for SWE-1.7. Those are vendor figures from the same post.

“Today we’re introducing SWE-2, our most advanced coding model yet. It pushes the Pareto frontier of capability and cost, achieving 50.0% on FrontierCode 1.1 Main, within one point of Fable 5.1 while being 64% cheaper.”The Cognition Team, Introducing SWE-2: Pushing the Pareto Frontier, September 10, 2026

What Alberta operators should do

Alberta work here is AI automation Alberta and AI consulting Alberta—choosing a coding-agent stack for shops that need repo edits, test loops, and review discipline without pretending a vendor bench is a purchase order. SWE-2’s availability trap matters more than the scoreboard: Cognition says it is available starting that day in Devin Desktop and CLI, rolling out on Devin Web and Fusion. The primary does not list a standalone API, open weights, or a per-token rate card. The model runs inside Devin.

Foreign Devin SaaS is not a private Alberta path. For client repos, credentials, and tender documents, prefer a private AI security boundary with human review before code or secrets leave the shop. If your team needs stack-selection training—when to use a hosted coding agent versus a private agent lane—that is AI consulting and AI training Edmonton work, not a blind paste of production code into a third-country endpoint.

The guardrail

A vendor claim that SWE-2 sits within one point of Fable 5.1 at 64% lower cost does not put inference in Alberta, and it does not authorize shipping regulated client code offshore without a written data boundary. For AI automation Alberta, use the Cognition primary as the scorecard, lock the Devin-only availability constraint into the RFP, and keep sensitive workloads on a private agents path. Do not invent a free-month promo—the SWE-2 primary we verified does not state one.

Opcelerate recommendationTreat SWE-2 as a Devin-hosted coding model with strong vendor cost–capability claims (50.0% FrontierCode 1.1 Main near Fable 5.1 at 64% cheaper; medium effort vs SWE-1.7: fewer turns, lower average cost). Lock the availability facts: Desktop + CLI now, Web/Fusion rolling out; no standalone API or open weights on the primary. Foreign SaaS ≠ private/on-prem. For Alberta shops, pair any Devin pilot with a human review gate and prefer private agents for client repos—Opcelerate can help with AI consulting Alberta, AI training Edmonton, and private AI security stack design.