Editorial re-verification: Devin Desktop
Cognition's Devin agentic software engineering product is priced from $0 for a Free plan with a "light quota to code with agents" up to $200/month for Max, with Pro at $20/month and Teams at $80/month…
Cognition's Devin agentic software engineering product is priced from $0 for a Free plan with a "light quota to code with agents" up to $200/month for Max, with Pro at $20/month and Teams at $80/month…
Cognition shipped macOS support and a Fusion local inference capability for its Devin coding agent in Devin Desktop and CLI, according to the company's blog. The releases move Devin beyond web-only sa…
Cognition gave its Devin coding agent a full macOS VM on bare-metal AWS hardware, a warmed-up Xcode, an iOS Simulator, and accessibility-tree vision so it can build, run, and verify iOS apps, then del…
Cognition's autonomous AI software engineering agent Devin is now available through AWS Marketplace, letting AWS enterprise customers deploy the agent for legacy modernization work including COBOL-to-…
A 2026 guide details how LLM model routing middleware can cut API costs by 40–85% without measurable quality loss by dispatching each request to the most suitable model. It outlines three architectura…
Cognition released SWE-2, a coding model post-trained with reinforcement learning from Moonshot AI's 2.8T-parameter Kimi K3, scoring 50.0% on FrontierCode 1.1 Main — within 1 point of Fable 5.1 at 64%…
Cognition released SWE-2, a coding model scoring 50% on FrontierCode 1.1 Main — within one point of Fable 5.1 — at 64% lower cost, built on a 2.8-trillion-parameter Kimi K3 base with RL post-training.…
Artificial Analysis benchmarked Cognition's Devin Fusion, which pairs a frontier model with the cheaper SWE-2 model as a sidekick, finding the GPT-6 Astra pairing costs 43 percent less and runs 31 per…
OpenAI published a case study on Cognition's Devin using GPT-6 Astra to autonomously test its own code, with one AI-generated summary claiming the model reduces engineers' manual review burden by 20–3…
Researchers at rubyhack.ai have linked a May 2026 attack on the RubyGems package repository to OpenAI agents that uploaded over 2,000 malicious packages attempting to steal API keys and execute arbitr…
Cognition's SWE-2 coding model scored 83.75% on the eight-task KingBench 3 benchmark, beating DeepSeek V4.1 Flash at 81.25% and its own base model Kimi K3 at 77.5%, according to independent testing. S…
Cognition released Fusion in Devin Desktop and CLI, a dual-agent harness that pairs a frontier "lead" model with a cheaper "sidekick" model and is up to 39% more efficient than other model harnesses a…
Cognition launched Devin Fusion, a coding agent architecture that pairs a frontier "lead" model for planning with a cheaper "sidekick" model for execution, delivering up to 39% better efficiency and a…
Shopify is rebuilding its mobile apps in Swift and Kotlin, six years after moving them to React Native, with its engineering blog crediting coding agents for cutting the cost of maintaining two platfo…
Cognition released SWE-2 on September 10, 2026, a coding model it says scores within one benchmark point of Anthropic's Fable 5.1 on FrontierCode 1.1 Main (50.0% vs 50.9%) at 64% lower compute cost, a…
Anthropic accused Alibaba, Moonshot, and DeepSeek of running large-scale distillation campaigns against Claude, alleging roughly 200 million exchanges were used to extract the model's reasoning traces…
A developer published a setup guide for a Python shim that exposes Cognition's SWE-2 and SWE-1.x coding models through an OpenAI-compatible API, letting tools like Claude Code call them via CLIProxyAP…
Dioxus Labs is joining Cognition to accelerate Devin, Cognition's autonomous cloud coding agent, according to a September 10, 2026 announcement by Dioxus creator Jonathan Kelley. The Dioxus team will …
Cognition released SWE-2, a 2.8T-parameter mixture-of-experts coding model built on the Kimi K3 base, which scored 92.8 on Terminal-Bench 2.1 — the highest figure in the published table — and 50.0 on …
Cognition released SWE-2 on September 10, a coding model it says approaches Fable 5.1's performance at 64% lower mean rollout cost, according to the company's own FrontierCode benchmark. SWE-2 scored …