agentboards.org
Compare/Factory Droid vs goose

Factory Droidvsgoose

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

Factory Droid
Factory · Terminal agent
#65MCP
Panel
6.4
0 spec wins
Reliability
6.5
Usefulness
7.0
Cost
5.7
Longevity
6.5

“Charges $200 a month for the Max plan and still governs you with rolling rate limits, so the ceiling is the product.”

goose
Agentic AI Foundation (originally Block) · Terminal agent
#13OSSMCP
Panel
6.8
5 spec wins
Reliability
6.3
Usefulness
6.5
Cost
6.8
Longevity
7.3

“Block gave it to the Linux Foundation, which is the one exit strategy that cannot be un-acquired.”

Spec by spec

SpecFactory Droidgoose
Architecture
CategoryTerminal agentTerminal agent
Runslocal, cloud, sandboxlocal
Platformsmacos, linux, windows, webmacos, linux, windows
Context windownot documentednot documented
Protocols
MCP clientYesYes
MCP serverNoYes
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsYesYes
Browser controlYesThe first-party Droid Control plugin (droid plugin install droid-control@factory-plugins) ships an `agent-browser` driver - a Playwright-backed CLI with Chrome DevTools Protocol support that navigates, fills forms, clicks and captures screenshots, powering /qa-test and /demo. YesThe built-in Computer Controller extension drives macOS apps including the browser through the Peekaboo CLI, rather than via a dedicated CDP or Playwright tool.
Sandboxed executionYesOS-level sandboxing isolates Droid from the filesystem and network using kernel-enforced policies (https://docs.factory.ai/autonomy-and-safety/sandbox.md). Yes`goose session|run --container <id>` executes extensions inside an existing Docker container; goose itself runs unsandboxed on the host otherwise.
Multi-agent orchestrationYesYes
Headless / CI modeYesYes
Models
BackboneClaude, GPT, Gemini, open-source and local models via BYOKClaude, GPT, Gemini, Ollama (local), OpenRouter, Azure OpenAI, Amazon Bedrock
Bring your own modelYesCustom models via `anthropic`, `openai` or `generic-chat-completion-api` providers with your own baseUrl and apiKey, plus an AWS Bedrock block with awsRegion/awsProfile. Yes
Local modelsYesBYOK config in ~/.factory/settings.json takes a baseUrl, so Ollama (http://localhost:11434/v1), LM Studio (http://localhost:1234/v1) and vLLM are all documented targets. Yes
Cost
Pricing modelsubscriptionbyok
Starts at$20/mo$0/mo
Free tierNoYes
Bring your own keyYesYes
Openness
Open sourceNoYes
LicenseproprietaryApache-2.0
GitHub stars4154,860

Which one would each critic pick

CriticFactory DroidgoosePick
El Juez——not enough reviews
El Amigo7.37.3no preference
El Crítico6.36.5goose — A local agent with terminal execution whose Docker isolation is a --container flag you have to remember, a foundation instead of a vendor, and a bill that depends entirely on which model you point it at.
El Profesor6.86.5Factory Droid — Droid's 58.8% Terminal-Bench result names the model and the leaderboard, which is more than most vendors manage, and its harness design is documented in hooks, MCP and Missions.
La Inversora7.05.3Factory Droid — Factory is selling to the org chart, with Slack, Teams, a Teams tier and Enterprise custom, and that is the buyer who tolerates rate limits, so the pricing ladder is a feature.
La Jefa6.55.8Factory Droid — Sixty seats on Teams is about $2,460 a month plus a rate limit nobody can budget, but the headless CI mode and Slack integration are the shape a team actually adopts.
El Hacker4.89.3goose — Rust, Apache-2.0, every extension an MCP server in YAML, Ollama on my box, a foundation instead of a vendor; this is the one I would build if I had the time.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.