agentboards.org
Compare/ArgusBot vs Wizard

ArgusBotvsWizard

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

ArgusBot
waltstephen · Agent harness
OSS
Panel
5.7
3 spec wins
Reliability
5.2
Usefulness
5.7
Cost
7.0
Longevity
5.0

“It ships a stall watchdog that restarts the agent, the most honest feature name anyone in this category has published.”

Wizard
teddytennant · Agent harness
OSS
Panel
5.7
1 spec wins
Reliability
5.0
Usefulness
5.5
Cost
7.3
Longevity
4.8

“It calls itself a sovereign agent, which is a great deal of constitutional language for a binary living in your home directory.”

Spec by spec

SpecArgusBotWizard
Architecture
CategoryAgent harnessAgent harness
Runslocallocal
Platformsmacos, linuxmacos, linux
Context windownot documentednot documented
Protocols
MCP clientNoNo
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesFiles are changed by the Codex or Claude Code backend ArgusBot supervises, not by ArgusBot itself. Yes
Git operationsNoNo
Browser controlNoNo
Sandboxed executionNoNo
Multi-agent orchestrationYesA reviewer sub-agent and a planner sub-agent run alongside the main agent to gate completion and maintain the plan. No
Headless / CI modeYesRuns can be started and steered without a terminal through a daemon controlled from Telegram or Feishu; there is no dedicated CI integration. No
Models
BackboneCodex CLI, Claude Code CLIxAI, OpenAI, Anthropic, Gemini, DeepSeek, Groq, OpenRouter, llama.cpp
Bring your own modelYesYes
Local modelsNoYes
Cost
Pricing modelbyokbyok
Starts at$0/mo$0/mo
Free tierYesYes
Bring your own keyYesYes
Openness
Open sourceYesYes
LicenseMITApache-2.0
GitHub stars317105

Which one would each critic pick

CriticArgusBotWizardPick
El Juez——not enough reviews
El Amigo6.36.0ArgusBot — Pick ArgusBot when your tasks have a test that says done; pick a plain agent session when done is a judgement call you would rather make yourself.
El Crítico5.34.8ArgusBot — Runs launched from the daemon use the permissive execution flag by default, and the README itself flags that as a security risk on untrusted workspaces.
El Profesor5.86.3Wizard — Memory is stored as markdown and sessions as line-delimited records, which makes the accumulated context a document a person can read and correct.
La Inversora5.55.0ArgusBot — MIT, one maintainer, 316 stars and no entity: this is a workflow opinion published as software, and workflow opinions do not have cap tables.
La Jefa4.84.0ArgusBot — Nothing to buy for sixty seats, and no way to govern them: the control surface is a Telegram or Feishu chat with slash commands and no directory behind it.
El Hacker6.88.0Wizard — Apache-2.0, and choosing Local on first run has it size a Qwen 3 build to my hardware and start llama.cpp itself, with no key and no account anywhere.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.