agentboards.org
Compare/Factory Droid vs Poolside

Factory DroidvsPoolside

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

Factory Droid
Factory · Terminal agent
#65MCP
Panel
6.4
1 spec wins
Reliability
6.5
Usefulness
7.0
Cost
5.7
Longevity
6.5

“Charges $200 a month for the Max plan and still governs you with rolling rate limits, so the ceiling is the product.”

Poolside
poolside · Terminal agent
#15MCP
Panel
6.9
1 spec wins
Reliability
6.8
Usefulness
7.0
Cost
6.5
Longevity
7.2

“One of its three network policies is named unsafe-allow-all, which is at least the most clearly labelled decision on this board.”

Spec by spec

SpecFactory DroidPoolside
Architecture
CategoryTerminal agentTerminal agent
Runslocal, cloud, sandboxlocal, sandbox, cloud
Platformsmacos, linux, windows, webmacos, linux, windows
Context windownot documented1M tokens on Laguna S 2.1, 256k on XS 2.1 and M.1
Protocols
MCP clientYesYes
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsYesYes
Browser controlYesThe first-party Droid Control plugin (droid plugin install droid-control@factory-plugins) ships an `agent-browser` driver - a Playwright-backed CLI with Chrome DevTools Protocol support that navigates, fills forms, clicks and captures screenshots, powering /qa-test and /demo. NoWeb search and fetch are documented; there is no DOM-level browser automation.
Sandboxed executionYesOS-level sandboxing isolates Droid from the filesystem and network using kernel-enforced policies (https://docs.factory.ai/autonomy-and-safety/sandbox.md). YesTool commands run inside a container on your machine and need a local Docker engine; network policy is off, allow-list or unsafe-allow-all, enforced through a proxy container.
Multi-agent orchestrationYesYes
Headless / CI modeYesYes`pool exec --prompt ... --output json` with POOLSIDE_API_KEY, plus a documented GitHub Actions integration.
Models
BackboneClaude, GPT, Gemini, open-source and local models via BYOKLaguna S 2.1, Laguna XS 2.1, Laguna M.1
Bring your own modelYesCustom models via `anthropic`, `openai` or `generic-chat-completion-api` providers with your own baseUrl and apiKey, plus an AWS Bedrock block with awsRegion/awsProfile. Yes
Local modelsYesBYOK config in ~/.factory/settings.json takes a baseUrl, so Ollama (http://localhost:11434/v1), LM Studio (http://localhost:1234/v1) and vLLM are all documented targets. YesDocumented local-run guides for Ollama and vLLM on Metal, plus full on-premises and air-gapped deployment.
Cost
Pricing modelsubscriptionmixed
Starts at$20/mon/a
Free tierNoYes
Bring your own keyYesYes
Openness
Open sourceNoNo
Licenseproprietaryproprietary
GitHub stars41n/a

Which one would each critic pick

CriticFactory DroidPoolsidePick
El Juez——not enough reviews
El Amigo7.36.8Factory Droid — Pick Droid if you want one agent in the terminal, in Slack and in CI; pick OpenCode if you would rather read the source and hold the keys yourself.
El Crítico6.36.5Poolside — Isolation needs a container engine running on the machine and enforces network policy through a proxy container, which is a configuration rather than a boundary.
El Profesor6.86.5Factory Droid — Droid's 58.8% Terminal-Bench result names the model and the leaderboard, which is more than most vendors manage, and its harness design is documented in hooks, MCP and Missions.
La Inversora7.07.0no preference
La Jefa6.57.0Poolside — It runs inside our own network, including air-gapped, on the Kubernetes platforms we already operate, and the enterprise price is a sales conversation.
El Hacker4.87.5Poolside — The client is proprietary and the model weights are published under OpenMDW-1.1 and Apache-2.0, which is the exact opposite of everyone else on this board.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.