Auggie CLIvsFactory Droid
Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.
Auggie CLI
Augment Code · Terminal agent
#43MCP
Panel
6.01 spec wins
- Reliability
- 6.2
- Usefulness
- 6.8
- Cost
- 5.3
- Longevity
- 5.5
“Windows support means Windows Subsystem for Linux, which is a polite way of saying Linux.”
Factory Droid
Factory · Terminal agent
#65MCP
Panel
6.45 spec wins
- Reliability
- 6.5
- Usefulness
- 7.0
- Cost
- 5.7
- Longevity
- 6.5
“Charges $200 a month for the Max plan and still governs you with rolling rate limits, so the ceiling is the product.”
Spec by spec
| Spec | Auggie CLI | Factory Droid |
|---|---|---|
| Architecture | ||
| Category | Terminal agent | Terminal agent |
| Runs | local | local, cloud, sandbox |
| Platforms | macos, linux, windows | macos, linux, windows, web |
| Context window | not documented | not documented |
| Protocols | ||
| MCP client | Yes | Yes |
| MCP server | Yes | No |
| Capabilities | ||
| Runs terminal commands | Yes | Yes |
| Multi-file edits | Yes | Yes |
| Git operations | YesAuggie analyses staged changes and diffs and supports a GitHub API token for repository operations. | Yes |
| Browser control | NoWeb search and web fetch tools are documented, but there is no browser automation. | YesThe first-party Droid Control plugin (droid plugin install droid-control@factory-plugins) ships an `agent-browser` driver - a Playwright-backed CLI with Chrome DevTools Protocol support that navigates, fills forms, clicks and captures screenshots, powering /qa-test and /demo. |
| Sandboxed execution | No | YesOS-level sandboxing isolates Droid from the filesystem and network using kernel-enforced policies (https://docs.factory.ai/autonomy-and-safety/sandbox.md). |
| Multi-agent orchestration | Yes | Yes |
| Headless / CI mode | Yes | Yes |
| Models | ||
| Backbone | Claude, Gemini, GPT-5, GLM, Kimi | Claude, GPT, Gemini, open-source and local models via BYOK |
| Bring your own model | No | YesCustom models via `anthropic`, `openai` or `generic-chat-completion-api` providers with your own baseUrl and apiKey, plus an AWS Bedrock block with awsRegion/awsProfile. |
| Local models | No | YesBYOK config in ~/.factory/settings.json takes a baseUrl, so Ollama (http://localhost:11434/v1), LM Studio (http://localhost:1234/v1) and vLLM are all documented targets. |
| Cost | ||
| Pricing model | subscription | subscription |
| Starts at | $20/mo | $20/mo |
| Free tier | No | No |
| Bring your own key | No | Yes |
| Openness | ||
| Open source | No | No |
| License | proprietary | proprietary |
| GitHub stars | n/a | 41 |
Which one would each critic pick
| Critic | Auggie CLI | Factory Droid | Pick |
|---|---|---|---|
| El Juez | — | — | not enough reviews |
| El Amigo | 6.8 | 7.3 | Factory Droid — Pick Droid if you want one agent in the terminal, in Slack and in CI; pick OpenCode if you would rather read the source and hold the keys yourself. |
| El Crítico | 5.8 | 6.3 | Factory Droid — The plans are priced in dollars and metered in rolling rate limits the pricing page does not publish, so the failure mode is being stopped, not being billed. |
| El Profesor | 6.0 | 6.8 | Factory Droid — Droid's 58.8% Terminal-Bench result names the model and the leaderboard, which is more than most vendors manage, and its harness design is documented in hooks, MCP and Missions. |
| La Inversora | 6.8 | 7.0 | Factory Droid — Factory is selling to the org chart, with Slack, Teams, a Teams tier and Enterprise custom, and that is the buyer who tolerates rate limits, so the pricing ladder is a feature. |
| La Jefa | 6.3 | 6.5 | Factory Droid — Sixty seats on Teams is about $2,460 a month plus a rate limit nobody can budget, but the headless CI mode and Slack integration are the shape a team actually adopts. |
| El Hacker | 4.3 | 4.8 | Factory Droid — Closed source, but BYOK reaches my Ollama box, MCP servers go in a JSON file and hooks are real, so Droid is the proprietary agent I resent the least. |
Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.