agentboards.org
Compare/Devika vs SWE-agent

DevikavsSWE-agent

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

Devika
Stition AI · Autonomous SWE
#250OSS
Panel
3.6
1 spec wins
Reliability
3.2
Usefulness
3.5
Cost
6.3
Longevity
1.3

“Nineteen thousand stars and nothing since 2025, which is the open-source version of a standing ovation on the way out.”

SWE-agent
Princeton University and Stanford University · Autonomous SWE
#144OSS
Panel
5.1
4 spec wins
Reliability
5.5
Usefulness
4.8
Cost
6.5
Longevity
3.7

“Maintenance-only and superseded by a version with mini in the name, which is the most honest changelog on this board.”

Spec by spec

SpecDevikaSWE-agent
Architecture
CategoryAutonomous SWEAutonomous SWE
Runslocallocal, sandbox
Platformsmacos, linux, windowsmacos, linux
Context windownot documentednot documented
Protocols
MCP clientNoNo
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsNoYes
Browser controlYesNo
Sandboxed executionNoYes
Multi-agent orchestrationNoNo
Headless / CI modeNoYes
Models
BackboneClaude, GPT, Gemini, MistralClaude, GPT, any LiteLLM-supported model
Bring your own modelYesYes
Local modelsYesYes
Cost
Pricing modelbyokbyok
Starts atn/a$0/mo
Free tierYesYes
Bring your own keyYesYes
Openness
Open sourceYesYes
LicenseMITMIT
GitHub stars19,55620,455

Which one would each critic pick

CriticDevikaSWE-agentPick
El Juez——not enough reviews
El Amigo3.55.0SWE-agent — SWE-agent is a research harness in maintenance mode, so read it, benchmark with it, and do not build your daily workflow on it; its own maintainers point you to mini-swe-agent.
El Crítico3.54.8SWE-agent — Maintenance-only by the maintainers' own statement, so what you are evaluating is a research artifact whose fixes now go to its successor.
El Profesor3.55.8SWE-agent — The agent-computer interface remains a principled design, its 2024 figure is properly scoped, and its newer results are described as leading without a number.
La Inversora3.54.0SWE-agent — There is no company here to outlive anything, only two universities and a successor project, so the 18-month question is about the fork tree, not the cap table.
La Jefa3.03.8SWE-agent — A maintenance-only research harness with no vendor, no contract and no SSO cannot pass a procurement review, however good its sandbox is.
El Hacker4.57.5SWE-agent — MIT, Docker-sandboxed, any LiteLLM model including my local server, and every prompt in the repo; maintenance-only just means the code stops changing under me.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.