no_humanvsSWE-agent
Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.
no_human
no_human · Autonomous SWE
#95OSSMCP
Panel
6.22 spec wins
- Reliability
- 5.7
- Usefulness
- 6.5
- Cost
- 6.8
- Longevity
- 5.8
“It runs an AI coding factory on your own machine, which is a stately way to describe the noise your laptop fan now makes.”
SWE-agent
Princeton University and Stanford University · Autonomous SWE
#144OSS
Panel
5.14 spec wins
- Reliability
- 5.5
- Usefulness
- 4.8
- Cost
- 6.5
- Longevity
- 3.7
“Maintenance-only and superseded by a version with mini in the name, which is the most honest changelog on this board.”
Spec by spec
| Spec | no_human | SWE-agent |
|---|---|---|
| Architecture | ||
| Category | Autonomous SWE | Autonomous SWE |
| Runs | local | local, sandbox |
| Platforms | macos, linux, windows | macos, linux |
| Context window | not documented | not documented |
| Protocols | ||
| MCP client | No | No |
| MCP server | Yes | No |
| Capabilities | ||
| Runs terminal commands | Yes | Yes |
| Multi-file edits | Yes | Yes |
| Git operations | Yes | Yes |
| Browser control | No | No |
| Sandboxed execution | No | Yes |
| Multi-agent orchestration | YesEach task runs a coder and then a separate adversarial reviewer model that never saw the coder's session, and several tasks run at once. | No |
| Headless / CI mode | No | Yes |
| Models | ||
| Backbone | any | Claude, GPT, any LiteLLM-supported model |
| Bring your own model | Yes | Yes |
| Local models | No | Yes |
| Cost | ||
| Pricing model | byok | byok |
| Starts at | $0/mo | $0/mo |
| Free tier | Yes | Yes |
| Bring your own key | Yes | Yes |
| Openness | ||
| Open source | Yes | Yes |
| License | MIT | MIT |
| GitHub stars | 328 | 20,455 |
Which one would each critic pick
| Critic | no_human | SWE-agent | Pick |
|---|---|---|---|
| El Juez | — | — | not enough reviews |
| El Amigo | 6.5 | 5.0 | no_human — Pick it if your repository has a test suite you trust; pick a plain coding agent if your tests are thin, because there is nothing here to catch that. |
| El Crítico | 5.5 | 4.8 | no_human — A reviewer told to refute completion will usually find something, and nothing in the row bounds the number of rounds before the argument stops. |
| El Profesor | 6.8 | 5.8 | no_human — The reviewer is a different model in a session that never saw the coder's work, instructed to refute completion rather than to assess it. |
| La Inversora | 6.0 | 4.0 | no_human — A marketing domain in front of a free tool that runs entirely on the user's hardware: it costs the maker nothing to operate and captures nothing either. |
| La Jefa | 5.3 | 3.8 | no_human — It does not run headless, so sixty developers each run the loop on a laptop under their own credentials and nothing opens a pull request I can attribute. |
| El Hacker | 7.3 | 7.5 | SWE-agent — MIT, Docker-sandboxed, any LiteLLM model including my local server, and every prompt in the repo; maintenance-only just means the code stops changing under me. |
Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.