agentboards.org
Compare/Crush vs Factory Droid

CrushvsFactory Droid

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

Crush
Charm · Terminal agent
#25OSSMCP
Panel
6.5
4 spec wins
Reliability
6.0
Usefulness
6.5
Cost
7.3
Longevity
6.2

“Has a --yolo flag, auto-discovers your local llama.cpp, and a license with MIT in the name that is not MIT yet.”

Factory Droid
Factory · Terminal agent
#65MCP
Panel
6.4
3 spec wins
Reliability
6.5
Usefulness
7.0
Cost
5.7
Longevity
6.5

“Charges $200 a month for the Max plan and still governs you with rolling rate limits, so the ceiling is the product.”

Spec by spec

SpecCrushFactory Droid
Architecture
CategoryTerminal agentTerminal agent
Runslocallocal, cloud, sandbox
Platformsmacos, linux, windowsmacos, linux, windows, web
Context windownot documentednot documented
Protocols
MCP clientYesYes
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsYesYes
Browser controlNoThe built-in tools stop at web_fetch, web_search and download; there is no browser, CDP or Playwright tool. YesThe first-party Droid Control plugin (droid plugin install droid-control@factory-plugins) ships an `agent-browser` driver - a Playwright-backed CLI with Chrome DevTools Protocol support that navigates, fills forms, clicks and captures screenshots, powering /qa-test and /demo.
Sandboxed executionNoShell and edit calls are guarded by permission prompts (or skipped with --yolo), with no container or OS-level isolation. YesOS-level sandboxing isolates Droid from the filesystem and network using kernel-enforced policies (https://docs.factory.ai/autonomy-and-safety/sandbox.md).
Multi-agent orchestrationNoYes
Headless / CI modeYes`crush run` executes a single non-interactive prompt, reads stdin and supports --quiet for scripts. Yes
Models
BackboneClaude, GPT, Gemini, any OpenAI-compatible or Anthropic-compatible providerClaude, GPT, Gemini, open-source and local models via BYOK
Bring your own modelYesYesCustom models via `anthropic`, `openai` or `generic-chat-completion-api` providers with your own baseUrl and apiKey, plus an AWS Bedrock block with awsRegion/awsProfile.
Local modelsYesOllama, LM Studio, llama.cpp and MLX providers are auto-discovered, and any OpenAI- or Anthropic-compatible base_url can be added. YesBYOK config in ~/.factory/settings.json takes a baseUrl, so Ollama (http://localhost:11434/v1), LM Studio (http://localhost:1234/v1) and vLLM are all documented targets.
Cost
Pricing modelbyoksubscription
Starts at$0/mo$20/mo
Free tierYesNo
Bring your own keyYesYes
Openness
Open sourceYesNo
LicenseFSL-1.1-MITproprietary
GitHub stars28,44441

Which one would each critic pick

CriticCrushFactory DroidPick
El Juez——not enough reviews
El Amigo7.37.3no preference
El Crítico6.36.3no preference
El Profesor6.56.8Factory Droid — Droid's 58.8% Terminal-Bench result names the model and the leaderboard, which is more than most vendors manage, and its harness design is documented in hooks, MCP and Missions.
La Inversora5.87.0Factory Droid — Factory is selling to the org chart, with Slack, Teams, a Teams tier and Enterprise custom, and that is the buyer who tolerates rate limits, so the pricing ladder is a feature.
La Jefa5.86.5Factory Droid — Sixty seats on Teams is about $2,460 a month plus a rate limit nobody can budget, but the headless CI mode and Slack integration are the shape a team actually adopts.
El Hacker7.54.8Crush — Reads my LSPs, finds my llama.cpp server on its own, MCP in a JSON file, a --yolo flag, and a license that keeps me from forking it today; close to great, grudgingly.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.