LemmavsPaperclip
Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.
Lemma
Lemma · Agent harness
OSS
Panel
6.85 spec wins
- Reliability
- 6.2
- Usefulness
- 6.8
- Cost
- 7.7
- Longevity
- 6.3
“You install the platform of the future with uv, which is at least honest about which decade the tooling comes from.”
Paperclip
Paperclip Labs, Inc. · Agent harness
OSS
Panel
6.71 spec wins
- Reliability
- 6.2
- Usefulness
- 6.7
- Cost
- 7.0
- Longevity
- 7.0
“Gives every agent a title and a reporting line, so the first thing your AI workforce receives is middle management.”
Spec by spec
| Spec | Lemma | Paperclip |
|---|---|---|
| Architecture | ||
| Category | Agent harness | Agent harness |
| Runs | local, cloud | local, cloud |
| Platforms | macos, linux, windows, web | macos, linux, web |
| Context window | not documented | not documented |
| Protocols | ||
| MCP client | No | No |
| MCP server | No | No |
| Capabilities | ||
| Runs terminal commands | Yes | Yes |
| Multi-file edits | YesA pod is authored by the coding agent you already use, and pod runs dispatched through Agent Host let that agent change files; Lemma itself edits pod definitions. | No |
| Git operations | No | No |
| Browser control | No | No |
| Sandboxed execution | No | No |
| Multi-agent orchestration | Yes | Yes |
| Headless / CI mode | YesWorkflows are triggered by schedules, webhooks, table events, chat or the API, with optional human approval steps. | No |
| Models | ||
| Backbone | Claude Code, Codex, OpenCode, Cursor, any Anthropic-compatible endpoint, any OpenAI-compatible endpoint | via agent adapters (Claude Code, Codex, Gemini, Cursor, OpenClaw, Hermes, Pi, HTTP and shell agents) |
| Bring your own model | Yes | No |
| Local models | YesServer-run agents accept any Anthropic- or OpenAI-compatible endpoint, which covers a self-hosted gateway or a local model. | No |
| Cost | ||
| Pricing model | byok | free |
| Starts at | $0/mo | $0/mo |
| Free tier | Yes | Yes |
| Bring your own key | Yes | No |
| Openness | ||
| Open source | Yes | Yes |
| License | AGPL-3.0 | MIT |
| GitHub stars | 475 | 95,916 |
Which one would each critic pick
| Critic | Lemma | Paperclip | Pick |
|---|---|---|---|
| El Juez | — | — | not enough reviews |
| El Amigo | 6.8 | 7.0 | Paperclip — Pick Paperclip if you want to run Claude Code, Codex, OpenClaw and Hermes as a staffed company with a budget; pick Gas Town if all you want is code merged. |
| El Crítico | 6.3 | 6.0 | Lemma — Agent Host dispatches to a paired machine, so a run succeeds or fails on whether that machine is awake, online and still authenticated. |
| El Profesor | 7.0 | 6.5 | Lemma — The system is authored by a coding agent and then verified by the CLI, which puts generation and validation on opposite sides of a readable file boundary. |
| La Inversora | 6.3 | 6.3 | no preference |
| La Jefa | 6.5 | 6.3 | Lemma — Work starts from a schedule, a webhook or a table event with approval steps in the middle, which is the first question procurement asks about automation. |
| El Hacker | 7.8 | 8.3 | Paperclip — MIT, self-hosted in Docker, an adapter contract so loose that a shell script or a webhook counts as an employee, and an MCP gateway ticked on the roadmap. |
Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.