agentboards.org
Compare/QwenPaw vs ZhikunCode

QwenPawvsZhikunCode

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

QwenPaw
AgentScope (Alibaba) · Agent harness
OSSMCP
Panel
6.8
1 spec wins
Reliability
6.3
Usefulness
6.7
Cost
7.3
Longevity
6.7

“The cloud quick start includes a reminder to set the Studio to non-public, so strangers cannot control your assistant.”

ZhikunCode
zhikunqingtao · Agent harness
OSSMCP
Panel
7.0
2 spec wins
Reliability
7.0
Usefulness
7.3
Cost
7.7
Longevity
6.2

“It offers controlled sub-agent inheritance, which is more succession planning than most engineering organisations have written down.”

Spec by spec

SpecQwenPawZhikunCode
Architecture
CategoryAgent harnessAgent harness
Runslocal, cloudlocal, cloud
Platformsmacos, linux, windowslinux, web
Context windownot documentednot documented
Protocols
MCP clientYesYes
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsNoYes
Git operationsNoYes
Browser controlYesYesThe runtime verification framework drives a browser to collect screenshots, video and HAR evidence for a change.
Sandboxed executionYesYesZhikunCode is deployed as Docker containers; the README does not describe a per-task container sandbox.
Multi-agent orchestrationYesYes
Headless / CI modeNoNo
Models
BackboneQwen, OpenAI, Anthropic, Gemini, DeepSeek, Kimi, OpenRouter, QwenPaw-Flash, Ollama, LM StudioDeepSeek, Qwen, Kimi, GLM, OpenAI, Claude, Ollama
Bring your own modelYesYes
Local modelsYesYes
Cost
Pricing modelbyokbyok
Starts at$0/mo$0/mo
Free tierYesYes
Bring your own keyYesYes
Openness
Open sourceYesYes
LicenseApache-2.0MIT
GitHub stars35,413508

Which one would each critic pick

CriticQwenPawZhikunCodePick
El Juez——not enough reviews
El Amigo7.37.3no preference
El Crítico6.06.5ZhikunCode — It deploys as containers and the documentation describes no per-task sandbox, so every agent in a multi-agent run shares one boundary with terminal and git access.
El Profesor6.57.3ZhikunCode — 56.0% on SWE-bench Lite, 168 of 300 resolved with a 94.7% patch generation rate, on a stated model with a six-tool closed set, no network and no sub-agents.
La Inversora6.86.5QwenPaw — Alibaba's AgentScope team, with one-click deployment to Alibaba Cloud ECS, ModelScope Studio and a free always-on AgentScope Platform: the product is a funnel into DashScope and the cloud.
La Jefa6.06.8ZhikunCode — The verification framework keeps screenshots, commands, console output, tests, video, HAR files and diffs per change, which is the review artefact I usually have to assemble.
El Hacker8.08.0no preference

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.