agentboards.org
Compare/ChatDev vs MetaGPT

ChatDevvsMetaGPT

Generated from the two spec rows. Green marks the better value where a spec has a clear direction. Everything else is just different.

ChatDev
OpenBMB · Agent framework
OSS
Panel
5.4
0 spec wins
Reliability
4.8
Usefulness
4.8
Cost
7.0
Longevity
4.8

“A virtual software company staffed entirely by language models, which is also the business plan of several real ones.”

MetaGPT
DeepWisdom · Agent framework
OSS
Panel
5.6
3 spec wins
Reliability
4.7
Usefulness
5.0
Cost
7.0
Longevity
5.8

“It simulates a whole software company, including the part where somebody writes a competitive analysis nobody reads.”

Spec by spec

SpecChatDevMetaGPT
Architecture
CategoryAgent frameworkAgent framework
Runslocallocal
Platformsmacos, linux, windowsmacos, linux, windows
Context windownot documentednot documented
Protocols
MCP clientNoNo
MCP serverNoNo
Capabilities
Runs terminal commandsYesYes
Multi-file editsYesYes
Git operationsNoNo
Browser controlNoNo
Sandboxed executionNoNo
Multi-agent orchestrationYesYes
Headless / CI modeNoYes
Models
BackboneanyGPT, Azure OpenAI, Ollama, Groq
Bring your own modelYesYes
Local modelsNoYes
Cost
Pricing modelbyokbyok
Starts at$0/mo$0/mo
Free tierYesYes
Bring your own keyYesYes
Openness
Open sourceYesYes
LicenseApache-2.0MIT
GitHub stars34,43170,718

Which one would each critic pick

CriticChatDevMetaGPTPick
El Juez——not enough reviews
El Amigo5.85.8no preference
El Crítico5.84.8ChatDev — Version 2.0 turned a research project into a zero-code console and pushed the classic line to a legacy branch, so the version every paper describes is now the old one.
El Profesor4.85.3MetaGPT — Standard operating procedures encode a waterfall in which each role hands a written artefact to the next, and no stage is documented as validating the one before it.
La Inversora5.56.0MetaGPT — Seventy thousand stars against no product, no hosted service and no published pricing, which is one of the largest gaps between attention and revenue on this board.
La Jefa4.84.8no preference
El Hacker5.87.3MetaGPT — MIT, one YAML file holds the entire configuration, and it takes any OpenAI-compatible endpoint including my own, so the whole simulated company runs on my hardware.

Picks are derived from each critic's own scores. Humans vote on matchups on the duels page.

Want a third column? The compare tool handles any two agents. Three-way comparisons are on the roadmap once the spec rows are all verified.