DevinvsSWE-agent
Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.
Devin
Cognition · Autonomous SWE
#139MCP
Panel
5.34 spec wins
- Reliability
- 5.3
- Usefulness
- 5.8
- Cost
- 4.3
- Longevity
- 5.5
“Replaced ACUs with credits in April 2026, so now the meter runs in a unit you already understand.”
SWE-agent
Princeton University and Stanford University · Autonomous SWE
#144OSS
Panel
5.15 spec wins
- Reliability
- 5.5
- Usefulness
- 4.8
- Cost
- 6.5
- Longevity
- 3.7
“Maintenance-only and superseded by a version with mini in the name, which is the most honest changelog on this board.”
Spec by spec
| Spec | Devin | SWE-agent |
|---|---|---|
| Architecture | ||
| Category | Autonomous SWE | Autonomous SWE |
| Runs | cloud, sandbox, local | local, sandbox |
| Platforms | web, macos, linux, windows | macos, linux |
| Context window | not documented | not documented |
| Protocols | ||
| MCP client | Yes | No |
| MCP server | Yes | No |
| Capabilities | ||
| Runs terminal commands | Yes | Yes |
| Multi-file edits | Yes | Yes |
| Git operations | Yes | Yes |
| Browser control | YesCloud sessions ship a first-party interactive Browser tool alongside the shell and IDE, with saved browser profiles for authenticated sites (https://docs.devin.ai/work-with-devin/browser-auth). | No |
| Sandboxed execution | YesCloud sessions run on a dedicated Devin machine built from environment blueprints and snapshots, and the Devin CLI adds OS-level isolation locally (https://docs.devin.ai/cli/sandbox). | Yes |
| Multi-agent orchestration | Yes | No |
| Headless / CI mode | Yes | Yes |
| Models | ||
| Backbone | Undisclosed (Cognition-managed models) | Claude, GPT, any LiteLLM-supported model |
| Bring your own model | NoA cross-provider model picker (`--model` / `/model` / config `agent.model`) selects between Anthropic, OpenAI, Google, Cognition and open-source models, but every one of them is served by Cognition: the CLI model reference documents no API key, base URL, gateway, Bedrock, Vertex or Azure deployment of your own, matching pricing.byok = false. | Yes |
| Local models | NoNo base URL, gateway or local endpoint setting exists in the Devin CLI config reference (https://docs.devin.ai/cli/reference/configuration/config-file). | Yes |
| Cost | ||
| Pricing model | mixed | byok |
| Starts at | $20/mo | $0/mo |
| Free tier | Yes | Yes |
| Bring your own key | NoUsage is billed in Cognition ACUs and no LLM provider key can be supplied; Devin Outposts (https://docs.devin.ai/cloud/outposts/overview) only moves session compute onto your own machines. | Yes |
| Openness | ||
| Open source | No | Yes |
| License | proprietary | MIT |
| GitHub stars | n/a | 20,455 |
Which one would each critic pick
| Critic | Devin | SWE-agent | Pick |
|---|---|---|---|
| El Juez | — | — | not enough reviews |
| El Amigo | 6.3 | 5.0 | Devin — The most complete autonomous stack here, with a cloud VM, browser and pull requests; pay for it if you can hand off whole tickets, not if you want a pair. |
| El Crítico | 5.3 | 4.8 | Devin — On-demand credits auto-refill, so an agent that loops spends money with nobody at the keyboard, and the meter is the risk on a tool built to run unattended. |
| El Profesor | 5.0 | 5.8 | SWE-agent — The agent-computer interface remains a principled design, its 2024 figure is properly scoped, and its newer results are described as leading without a number. |
| La Inversora | 6.8 | 4.0 | Devin — Well funded, acquisitive, and repricing in public; the Windsurf purchase bought distribution, and the April 2026 plan change says the ACU math was not working. |
| La Jefa | 5.8 | 3.8 | Devin — Teams at $80 plus $40 per full seat with shared credits, and Enterprise by quote; the free flex seats are the part finance will like. |
| El Hacker | 2.5 | 7.5 | SWE-agent — MIT, Docker-sandboxed, any LiteLLM model including my local server, and every prompt in the repo; maintenance-only just means the code stops changing under me. |
Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.