agentboards.org
Compare/Herm vs mini-SWE-agent

Hermvsmini-SWE-agent

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

Herm
aduermael · Terminal agent
#102OSS
Panel
7.2
1 spec wins
Reliability
7.2
Usefulness
6.8
Cost
8.5
Longevity
6.2

“It offers an in-process Unix-like sandbox, for the days when Docker feels like too much of a commitment.”

mini-SWE-agent
SWE-agent (Princeton and Stanford) · Terminal agent
#35OSS
Panel
6.9
3 spec wins
Reliability
6.3
Usefulness
5.5
Cost
8.8
Longevity
7.0

“One hundred lines of Python, which is fewer than most competitors spend on their pricing page.”

Spec by spec

SpecHermmini-SWE-agent
Architecture
CategoryTerminal agentTerminal agent
Runslocal, sandboxlocal, sandbox
Platformsmacos, linuxmacos, linux
Context windownot documentednot documented
Protocols
MCP clientNoNo
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsNoYes
Browser controlNoNo
Sandboxed executionYesYes
Multi-agent orchestrationYesA session can mix models by role, such as one model as the main agent, a cheaper one for exploration and another for vision. No
Headless / CI modeNoYes
Models
BackboneAnthropic, OpenAI, Gemini, Grok, OpenRouter, Ollama, Azure OpenAI, Vertex AI, AWS Bedrockany via litellm
Bring your own modelYesYes
Local modelsYesYes
Cost
Pricing modelbyokbyok
Starts at$0/mon/a
Free tierYesYes
Bring your own keyYesYes
Openness
Open sourceYesYes
LicenseMITMIT
GitHub stars2348,145

Which one would each critic pick

CriticHermmini-SWE-agentPick
El Juez——not enough reviews
El Amigo7.56.8Herm — Pick it if approval prompts have trained you to click yes without reading; pick a host-based agent if installing Docker is a fight you would rather skip.
El Crítico6.86.5Herm — It extends its own environment by writing Dockerfiles dynamically, which means the agent authors the definition of the box it is confined to.
El Profesor7.88.0mini-SWE-agent — The reference harness for the SWE-bench bash-only leaderboard: one bash action per turn via subprocess.run, a linear history, no tool-calling API, and Gemini 3 Pro reported above 74% on Verified.
La Inversora6.06.8mini-SWE-agent — No company, a Princeton and Stanford lab with 6,938 stars; the funding is grants and the exit is a paper, which is more stable than half the cap tables on this board.
La Jefa6.85.0Herm — It defaults to running the agent in Docker, which means a container runtime on sixty machines, and a Homebrew tap is the closest thing here to a distribution channel.
El Hacker8.38.5mini-SWE-agent — MIT, any model through litellm, OpenRouter or Portkey including my local server, a YAML config, and source short enough to read before breakfast; the missing MCP client is the only gap.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.