agentboards.org
Compare/codehamr vs pi

codehamrvspi

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

codehamr
codehamr · Terminal agent
#115OSS
Panel
6.7
0 spec wins
Reliability
6.3
Usefulness
5.8
Cost
9.3
Longevity
5.2

“Four tools and three slash commands, which is fewer features than most agents put in the onboarding screen.”

pi
Earendil Inc. · Terminal agent
#9OSS
Panel
6.8
2 spec wins
Reliability
6.5
Usefulness
6.5
Cost
7.3
Longevity
7.0

“A coding agent that ships themes, which is the first time anyone has cared what the terminal looks like while a file gets deleted.”

Spec by spec

Speccodehamrpi
Architecture
CategoryTerminal agentTerminal agent
Runslocallocal
Platformsmacos, linux, windowsmacos, linux, windows
Context windownot documentednot documented
Protocols
MCP clientNoNo
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsNoNo
Browser controlNoThe agent can bootstrap a headless browser as a verification helper when the machine allows it, but it has no browser tool of its own. No
Sandboxed executionNoNo
Multi-agent orchestrationNoNo
Headless / CI modeNoYes
Models
BackboneOllama, vLLM, LM Studio, OpenAI-compatible endpointsOpenAI, Anthropic, Google, OpenRouter, llama.cpp
Bring your own modelYesYes
Local modelsYesYes
Cost
Pricing modelbyokbyok
Starts at$0/mo$0/mo
Free tierYesYes
Bring your own keyYesYes
Openness
Open sourceYesYes
LicenseMITMIT
GitHub stars219111,339

Which one would each critic pick

CriticcodehamrpiPick
El Juez——not enough reviews
El Amigo6.87.5pi — Pick pi if you want a terminal agent small enough to read in an afternoon and extend in TypeScript; pick OpenCode if your workflow already depends on MCP servers.
El Crítico6.36.8pi — Extensions are TypeScript modules loaded into the agent's own process, so a package is a supply chain, and the install line already says --ignore-scripts.
El Profesor7.87.3codehamr — It runs the tests, compiles, or loads the page, and prints `unverified:` when it cannot, which is the rarest honesty in this category.
La Inversora6.05.5codehamr — 216 stars, and the only commercial artefact is a hosted endpoint sitting behind a waitlist, which is a business plan in its earliest possible form.
La Jefa5.85.5codehamr — Nothing per seat, nothing in the pipeline, no console and no audit log, and the Windows contingent needs WSL2 before they can start at all.
El Hacker7.58.5pi — MIT TypeScript, extensions that add tools, commands and UI, /llama for llama.cpp on my box, ANTHROPIC_API_KEY in the environment, and packages to share it all.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.