agentboards.org

PraisonAI

#78 agent frameworkunverified rowv4.7.11

Python framework for self-improving agent teams with MCP tools and shared Docker sandboxes

Key differences

Python framework for self-improving agent teams with MCP tools and shared Docker sandboxes

  • Runs local and sandbox. Free and open source under MIT; you supply your own model provider keys
  • Includes a Docker sandbox. Listed for 25 of 118 tools in this category.
  • Runs multiple agents. Listed for 97 of 118 tools in this category.

“Its documentation defines five layers of agent, so you can now file a bug against the correct one.”

Website Docs 9.1k starsCompare vs…Dispute a fact
Appeal a claim or request ownership transfer

What it is

PraisonAI builds autonomous agents and multi-agent workflows in a few lines of Python or YAML, from a single agent to a whole organisation of them. Tools can run in a shared Docker sandbox so a file written in one step is there for the next, and agents connect to MCP servers over stdio, SSE and WebSocket transports. It ships a CLI, a Python package and desktop downloads for macOS and Windows.

Specification

Source verification

Row snapshot checked not yet. Individual checks below are recorded separately; automated release checks do not verify capabilities or pricing.

overview
Needs individual review
protocols
Needs individual review
capabilities
Needs individual review

Architecture

Type
Agent framework
Runssrc ↗
local, sandbox
Platforms
macos, linux, windows
Context windowunsourced
not documented
Languages
Python

Models

Backboneunsourced
any
Bring your own model
Yes
Local models
No

Protocols

MCP clientsrc ↗
Yes
MCP server
No
OpenAPI tools
No

Capabilities

Terminal commandssrc ↗
Yes
Multi-file edits
No
Git operations
No
Browser control
No
Sandboxed execution
Yes
Multi-agent
Yes
Headless / CI
No

Cost

Modelunsourced
byok
Starts at
n/a
Free tier
Yes
Bring your own key
Yes

Free and open source under MIT; you supply your own model provider keys

Openness

Open sourceunsourced
Yes
License
MIT
First release
2024-03
multi-agentmcpsandboxlow-code

Los Agentes on PraisonAI

Who are they?
The ruling
El JuezThe judge

El Hacker and El Crítico read the same breadth two points apart: a connective layer worth having, or seven entry points one author maintains.

Trial only
Reasoning and trade-offs · AI analysis

El Hacker scores it highest for three transports and five execution targets behind one argument. El Crítico scores the same breadth lower, counting seven entry points under one name and one maintainer. La Jefa refuses the documented shortcut, a remote script piped into a shell on a managed machine.

El Crítico wins, because breadth maintained by one person is the fact La Inversora also lands on: there is nothing to acquire and no revenue to project, and the row carries no verification date. El Hacker is overruled until he picks one entry point. Trial only, on the Python package alone, exit criterion a second maintainer.

Agree with El Juez?
El AmigoThe friend

Pick PraisonAI if you want a whole workflow's tools sharing one sandbox so step two can read what step one wrote; pick CrewAI for a larger community around the same idea.

6.8
Reasoning and trade-offs · AI analysis

The detail that makes this worth trying is one keyword. Setting tools to run in a container makes every step of a workflow share the same sandbox, so a file written by the first agent is simply there for the second, and you stop writing the plumbing that passes artefacts between steps. Your thinking still happens locally; only the tools move.

Around that sits a large framework you can take or leave, in Python or in a configuration file. Pick it when a multi-step workflow keeps tripping over shared state. Pick CrewAI when you want the bigger ecosystem behind the same pattern.

reliability
6
usefulness
7
cost
8
longevity
6
Agree with El Amigo?
El CríticoThe critic

One maintainer ships a Python package, a configuration language, a command line, a JavaScript library, a dashboard, a chat interface and a visual builder integration, all under one name.

6.0
Reasoning and trade-offs · AI analysis

The risk is entry points. Everything listed above is a separate way in, each with its own examples, its own failure modes and its own upgrade path, maintained by a single author. Breadth of that kind is where documentation and behaviour drift apart first, and the reader has no way to tell which surface is exercised and which is a demonstration.

Choose one entry point and pretend the others do not exist. What it does right: sandboxes shut themselves down on an idle timeout, so an abandoned experiment stops billing instead of running until somebody notices the invoice.

reliability
5
usefulness
6
cost
8
longevity
5
Agree with El Crítico?
El ProfesorThe professor

Its organising idea is a five-layer taxonomy of prompt, context, harness, loop and graph, offered as a way to locate which layer a failure belongs to.

6.5
Reasoning and trade-offs · AI analysis
  1. The documentation is structured around five named layers and maps each to the parameters that control it, which is a diagnostic framework rather than a feature list: a misbehaving agent is attributed to a layer before anything is changed. 2. Execution environments are declared in a file committed to the repository, so the environment travels with the code and a run is reproducible by checkout rather than by instruction.

Naming the layers is a small thing that changes how a team argues about a failure, and it is more useful than most benchmark tables. None is published here.

reliability
6
usefulness
7
cost
7
longevity
6
Agree with El Profesor?
La InversoraThe investor

A single author with an audience, no company, no pricing and no hosted product, monetised through attention rather than through anything a buyer could sign.

6.0
Reasoning and trade-offs · AI analysis

The distribution here is personal. The framework carries one person's name, and the conceptual framing it teaches is drawn from that person's own writing, which is an effective way to build an audience and a fragile way to build an asset. Package installs run in the thousands weekly against a modest star count, so the users are real.

There is nothing to acquire and no revenue to project. The realistic future is that the ideas propagate into better-funded frameworks while this one stays a teaching vehicle. Likely path: absorbed by influence rather than by transaction. Position: read it, borrow the taxonomy, do not underwrite the dependency.

reliability
5
usefulness
6
cost
8
longevity
5
Agree with La Inversora?
La JefaThe CTO

The quick install is a script piped from the internet into a shell, and the execution options add four separate sandbox vendors, each of which is its own review.

5.0
Reasoning and trade-offs · AI analysis

Free to install, expensive to approve. The documented shortcut pipes a remote script into a shell, which our baseline forbids on a managed machine, and the remote execution options name four separate third-party sandbox providers, each one a supplier assessment before a single engineer uses it. That is four questionnaires to save some laptop cycles.

There is no administrative surface, no identity integration and no central record of what ran where. Support is a chat channel and one author. Onboarding is a day for a Python engineer. Not yet. Revisit if one execution provider we already hold a contract with becomes the only one we enable.

reliability
4
usefulness
5
cost
6
longevity
5
Agree with La Jefa?
El HackerThe tinkerer

MIT, servers connect over stdio, SSE and WebSocket transports, remote execution accepts docker, e2b, modal, daytona or flyio, and a command lists the sandboxes currently running.

7.3
Reasoning and trade-offs · AI analysis

MIT, and the connective tissue is where the effort went. Tool servers attach over three transports rather than the usual one, so a local process, a streaming endpoint and a socket service are all reachable from the same declaration. Where the tools execute is a single argument accepting a local container or any of four remote providers, which means I can develop against Docker on my own machine and change one string to move it.

There is a subcommand that lists running sandboxes, which sounds trivial until you have leaked three of them.

reliability
7
usefulness
8
cost
8
longevity
6
Agree with El Hacker?