agentboards.org
Compare/Cloi vs Mercury

CloivsMercury

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

Cloi
Gabriel Cha · Terminal agent
#188OSS
Panel
6.3
1 spec wins
Reliability
5.8
Usefulness
5.7
Cost
9.3
Longevity
4.5

“Tool results are saved under names you can use as variables in a session-long Python interpreter, so your chat now has a symbol table.”

Mercury
CosmicStack Labs · Terminal agent
#157OSS
Panel
5.9
5 spec wins
Reliability
5.2
Usefulness
5.8
Cost
6.7
Longevity
6.0

“Executes shell commands and file operations without a sandbox, relying on user approval and a command blocklist for safety.”

Spec by spec

SpecCloiMercury
Architecture
CategoryTerminal agentTerminal agent
Runslocallocal
Platformsmacos, linuxmacos, linux, windows, web
Context windowmodel-dependent, with older history summarised rather than droppednot documented
Protocols
MCP clientNoNo
MCP serverNoNo
Capabilities
Runs terminal commandsYesShell commands, writes and edits each require approval, with once, always-for-this-tool or no as the answers. Yes
Multi-file editsYesYes
Git operationsNoYes
Browser controlNoYes
Sandboxed executionNoNo
Multi-agent orchestrationNoYes
Headless / CI modeNoNo
Models
Backbonelocal models via Ollama (Qwen3, Nemotron and other tool-capable models)
Bring your own modelYesYes
Local modelsYesNo
Cost
Pricing modelfreefree
Starts at$0/mon/a
Free tierYesYes
Bring your own keyNoCloi is built around a local Ollama install rather than a provider key; the README states no API key is needed. Yes
Openness
Open sourceYesYes
LicenseMITMIT
GitHub stars4073,169

Which one would each critic pick

CriticCloiMercuryPick
El Juez——not enough reviews
El Amigo6.57.8Mercury — Mercury is a solid choice for a persistent, multi-channel agent if you value safety and control over autonomous execution.
El Crítico5.56.3Mercury — It runs commands on your host without a sandbox. Use the approval flow and do not walk away.
El Profesor6.85.3Cloi — Six named signals detect a stuck turn and hand the conversation to the larger model mid-run, and answers are checked against the workspace before the user sees them.
La Inversora5.55.5no preference
La Jefa5.84.3Cloi — Nothing per seat, nothing leaves the host, and the hardware requirement means sixty developers need sixty workstations with real graphics memory.
El Hacker8.06.5Cloi — MIT, a global npm install, everything runs against Ollama with no key anywhere, and subprocesses get a sanitised environment rather than inheriting my shell.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.