AI ReviewvsShippie
Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.
AI Review
Nikita Filonov · Code review agent
#52OSS
Panel
7.01 spec wins
- Reliability
- 6.8
- Usefulness
- 6.8
- Cost
- 8.0
- Longevity
- 6.2
“It replies inside your existing review threads, so the argument you lost in March can now continue without you.”
Shippie
Matt Carey · Code review agent
#177OSSMCP
Panel
6.33 spec wins
- Reliability
- 6.0
- Usefulness
- 6.2
- Cost
- 8.0
- Longevity
- 5.2
“You summon the reviewer by typing /shippie review, which is the first code review process anyone has ever voluntarily started.”
Spec by spec
| Spec | AI Review | Shippie |
|---|---|---|
| Architecture | ||
| Category | Code review agent | Code review agent |
| Runs | local, cloud | local, cloud |
| Platforms | macos, linux, windows | macos, linux, web |
| Context window | not documented | not documented |
| Protocols | ||
| MCP client | No | Yes |
| MCP server | No | No |
| Capabilities | ||
| Runs terminal commands | YesOnly in agent mode, where the model runs read-only shell commands such as ls, cat, rg and git to explore the repository before reviewing. | Yes |
| Multi-file edits | No | No |
| Git operations | YesIt reads diffs from and writes comments and thread replies to six hosted VCS platforms; it does not commit or push. | Yes |
| Browser control | No | No |
| Sandboxed execution | No | No |
| Multi-agent orchestration | No | Yes |
| Headless / CI mode | Yes | Yes |
| Models | ||
| Backbone | OpenAI, Claude, Gemini, Ollama, AWS Bedrock, OpenRouter, Azure OpenAI | Anthropic, OpenAI, OpenRouter, Cloudflare Workers AI |
| Bring your own model | Yes | Yes |
| Local models | YesOllama is a first-class provider. | No |
| Cost | ||
| Pricing model | byok | byok |
| Starts at | $0/mo | $0/mo |
| Free tier | Yes | Yes |
| Bring your own key | Yes | Yes |
| Openness | ||
| Open source | Yes | Yes |
| License | Apache-2.0 | MIT |
| GitHub stars | 585 | 2,502 |
Which one would each critic pick
| Critic | AI Review | Shippie | Pick |
|---|---|---|---|
| El Juez | — | — | not enough reviews |
| El Amigo | 7.3 | 6.3 | AI Review — Pick it if your code lives on Gitea or Azure DevOps and every review bot you tried only speaks GitHub; pick a hosted reviewer if GitHub is all you have. |
| El Crítico | 6.3 | 6.0 | AI Review — Agent mode is an iterative loop in which the model runs shell commands until it decides to stop, and no documented ceiling bounds the turns or the tokens they consume. |
| El Profesor | 7.0 | 6.3 | AI Review — Inline, context and summary prompts are all user-supplied, so the output distribution is a property of a team's configuration rather than of the tool being evaluated. |
| La Inversora | 6.3 | 6.0 | AI Review — 565 stars, one author and a package on PyPI: the whole business is somebody's evenings, and every platform it integrates with is building this feature in-house. |
| La Jefa | 7.3 | 6.0 | AI Review — It runs inside the CI job we already pay for, with no seat licence and no vendor holding our diffs, which is the shortest procurement conversation of the quarter. |
| El Hacker | 7.8 | 7.5 | AI Review — Apache-2.0, Ollama as a first-class provider, and config from YAML, JSON or environment variables, so a fully local reviewer is a file I write rather than a plan I request. |
Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.