agentboards.org
Compare/Devin vs SWE-agent

DevinvsSWE-agent

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

Devin
Cognition · Autonomous SWE
#139MCP
Panel
5.3
4 spec wins
Reliability
5.3
Usefulness
5.8
Cost
4.3
Longevity
5.5

“Replaced ACUs with credits in April 2026, so now the meter runs in a unit you already understand.”

SWE-agent
Princeton University and Stanford University · Autonomous SWE
#144OSS
Panel
5.1
5 spec wins
Reliability
5.5
Usefulness
4.8
Cost
6.5
Longevity
3.7

“Maintenance-only and superseded by a version with mini in the name, which is the most honest changelog on this board.”

Spec by spec

SpecDevinSWE-agent
Architecture
CategoryAutonomous SWEAutonomous SWE
Runscloud, sandbox, locallocal, sandbox
Platformsweb, macos, linux, windowsmacos, linux
Context windownot documentednot documented
Protocols
MCP clientYesNo
MCP serverYesNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsYesYes
Browser controlYesCloud sessions ship a first-party interactive Browser tool alongside the shell and IDE, with saved browser profiles for authenticated sites (https://docs.devin.ai/work-with-devin/browser-auth). No
Sandboxed executionYesCloud sessions run on a dedicated Devin machine built from environment blueprints and snapshots, and the Devin CLI adds OS-level isolation locally (https://docs.devin.ai/cli/sandbox). Yes
Multi-agent orchestrationYesNo
Headless / CI modeYesYes
Models
BackboneUndisclosed (Cognition-managed models)Claude, GPT, any LiteLLM-supported model
Bring your own modelNoA cross-provider model picker (`--model` / `/model` / config `agent.model`) selects between Anthropic, OpenAI, Google, Cognition and open-source models, but every one of them is served by Cognition: the CLI model reference documents no API key, base URL, gateway, Bedrock, Vertex or Azure deployment of your own, matching pricing.byok = false. Yes
Local modelsNoNo base URL, gateway or local endpoint setting exists in the Devin CLI config reference (https://docs.devin.ai/cli/reference/configuration/config-file). Yes
Cost
Pricing modelmixedbyok
Starts at$20/mo$0/mo
Free tierYesYes
Bring your own keyNoUsage is billed in Cognition ACUs and no LLM provider key can be supplied; Devin Outposts (https://docs.devin.ai/cloud/outposts/overview) only moves session compute onto your own machines. Yes
Openness
Open sourceNoYes
LicenseproprietaryMIT
GitHub starsn/a20,455

Which one would each critic pick

CriticDevinSWE-agentPick
El Juez——not enough reviews
El Amigo6.35.0Devin — The most complete autonomous stack here, with a cloud VM, browser and pull requests; pay for it if you can hand off whole tickets, not if you want a pair.
El Crítico5.34.8Devin — On-demand credits auto-refill, so an agent that loops spends money with nobody at the keyboard, and the meter is the risk on a tool built to run unattended.
El Profesor5.05.8SWE-agent — The agent-computer interface remains a principled design, its 2024 figure is properly scoped, and its newer results are described as leading without a number.
La Inversora6.84.0Devin — Well funded, acquisitive, and repricing in public; the Windsurf purchase bought distribution, and the April 2026 plan change says the ACU math was not working.
La Jefa5.83.8Devin — Teams at $80 plus $40 per full seat with shared credits, and Enterprise by quote; the free flex seats are the part finance will like.
El Hacker2.57.5SWE-agent — MIT, Docker-sandboxed, any LiteLLM model including my local server, and every prompt in the repo; maintenance-only just means the code stops changing under me.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.