CrushvsFactory Droid
Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.
Crush
Charm · Terminal agent
#25OSSMCP
Panel
6.54 spec wins
- Reliability
- 6.0
- Usefulness
- 6.5
- Cost
- 7.3
- Longevity
- 6.2
“Has a --yolo flag, auto-discovers your local llama.cpp, and a license with MIT in the name that is not MIT yet.”
Factory Droid
Factory · Terminal agent
#65MCP
Panel
6.43 spec wins
- Reliability
- 6.5
- Usefulness
- 7.0
- Cost
- 5.7
- Longevity
- 6.5
“Charges $200 a month for the Max plan and still governs you with rolling rate limits, so the ceiling is the product.”
Spec by spec
| Spec | Crush | Factory Droid |
|---|---|---|
| Architecture | ||
| Category | Terminal agent | Terminal agent |
| Runs | local | local, cloud, sandbox |
| Platforms | macos, linux, windows | macos, linux, windows, web |
| Context window | not documented | not documented |
| Protocols | ||
| MCP client | Yes | Yes |
| MCP server | No | No |
| Capabilities | ||
| Runs terminal commands | Yes | Yes |
| Multi-file edits | Yes | Yes |
| Git operations | Yes | Yes |
| Browser control | NoThe built-in tools stop at web_fetch, web_search and download; there is no browser, CDP or Playwright tool. | YesThe first-party Droid Control plugin (droid plugin install droid-control@factory-plugins) ships an `agent-browser` driver - a Playwright-backed CLI with Chrome DevTools Protocol support that navigates, fills forms, clicks and captures screenshots, powering /qa-test and /demo. |
| Sandboxed execution | NoShell and edit calls are guarded by permission prompts (or skipped with --yolo), with no container or OS-level isolation. | YesOS-level sandboxing isolates Droid from the filesystem and network using kernel-enforced policies (https://docs.factory.ai/autonomy-and-safety/sandbox.md). |
| Multi-agent orchestration | No | Yes |
| Headless / CI mode | Yes`crush run` executes a single non-interactive prompt, reads stdin and supports --quiet for scripts. | Yes |
| Models | ||
| Backbone | Claude, GPT, Gemini, any OpenAI-compatible or Anthropic-compatible provider | Claude, GPT, Gemini, open-source and local models via BYOK |
| Bring your own model | Yes | YesCustom models via `anthropic`, `openai` or `generic-chat-completion-api` providers with your own baseUrl and apiKey, plus an AWS Bedrock block with awsRegion/awsProfile. |
| Local models | YesOllama, LM Studio, llama.cpp and MLX providers are auto-discovered, and any OpenAI- or Anthropic-compatible base_url can be added. | YesBYOK config in ~/.factory/settings.json takes a baseUrl, so Ollama (http://localhost:11434/v1), LM Studio (http://localhost:1234/v1) and vLLM are all documented targets. |
| Cost | ||
| Pricing model | byok | subscription |
| Starts at | $0/mo | $20/mo |
| Free tier | Yes | No |
| Bring your own key | Yes | Yes |
| Openness | ||
| Open source | Yes | No |
| License | FSL-1.1-MIT | proprietary |
| GitHub stars | 28,444 | 41 |
Which one would each critic pick
| Critic | Crush | Factory Droid | Pick |
|---|---|---|---|
| El Juez | — | — | not enough reviews |
| El Amigo | 7.3 | 7.3 | no preference |
| El Crítico | 6.3 | 6.3 | no preference |
| El Profesor | 6.5 | 6.8 | Factory Droid — Droid's 58.8% Terminal-Bench result names the model and the leaderboard, which is more than most vendors manage, and its harness design is documented in hooks, MCP and Missions. |
| La Inversora | 5.8 | 7.0 | Factory Droid — Factory is selling to the org chart, with Slack, Teams, a Teams tier and Enterprise custom, and that is the buyer who tolerates rate limits, so the pricing ladder is a feature. |
| La Jefa | 5.8 | 6.5 | Factory Droid — Sixty seats on Teams is about $2,460 a month plus a rate limit nobody can budget, but the headless CI mode and Slack integration are the shape a team actually adopts. |
| El Hacker | 7.5 | 4.8 | Crush — Reads my LSPs, finds my llama.cpp server on its own, MCP in a JSON file, a --yolo flag, and a license that keeps me from forking it today; close to great, grudgingly. |
Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.