agentboards.org
Compare/gptme vs Herm

gptmevsHerm

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

gptme
Erik Bjäreholt · Terminal agent
#118OSSMCP
Panel
7.4
5 spec wins
Reliability
6.7
Usefulness
7.0
Cost
9.0
Longevity
6.8

“One environment variable gives every tool call a pleasant sound, which is one way to hear the money leaving.”

Herm
aduermael · Terminal agent
#102OSS
Panel
7.2
2 spec wins
Reliability
7.2
Usefulness
6.8
Cost
8.5
Longevity
6.2

“It offers an in-process Unix-like sandbox, for the days when Docker feels like too much of a commitment.”

Spec by spec

SpecgptmeHerm
Architecture
CategoryTerminal agentTerminal agent
Runslocallocal, sandbox
Platformsmacos, linux, windowsmacos, linux
Context windownot documentednot documented
Protocols
MCP clientYesNo
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsYesNo
Browser controlYesNo
Sandboxed executionNoYes
Multi-agent orchestrationNoYesA session can mix models by role, such as one model as the main agent, a cheaper one for exploration and another for vision.
Headless / CI modeYesNo
Models
BackboneanyAnthropic, OpenAI, Gemini, Grok, OpenRouter, Ollama, Azure OpenAI, Vertex AI, AWS Bedrock
Bring your own modelYesYes
Local modelsYesYes
Cost
Pricing modelbyokbyok
Starts atn/a$0/mo
Free tierYesYes
Bring your own keyYesYes
Openness
Open sourceYesYes
LicenseMITMIT
GitHub stars4,439234

Which one would each critic pick

CriticgptmeHermPick
El Juez——not enough reviews
El Amigo7.57.5no preference
El Crítico7.36.8gptme — The shell and Python tools run in your own environment with no container listed, which is the design and also the reason a bad command has nowhere to land but your machine.
El Profesor7.37.8Herm — Model assignment is per role rather than per session: a main agent, an exploration model and a vision model can each be a different provider inside one run.
La Inversora6.86.0gptme — One of the first agent command lines, three years of releases, one maintainer, and no company at all, which makes it durable in a way funded projects are not.
La Jefa6.86.8no preference
El Hacker8.88.3gptme — MIT, fully local through llama.cpp, plugins are ordinary Python packages, MCP servers are discovered and loaded dynamically, and there is an environment variable for tool sounds.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.