agentboards.org

AutoGPT

#42 agent harnessverified Sep 3, 2026autogpt-platform-beta-v0.8.2

The original autonomous-agent project, now a platform to build, schedule and run agent workflows

Key differences

The original autonomous-agent project, now a platform to build, schedule and run agent workflows

  • Runs local and cloud. Self-hosting is free with your own API keys; the managed platform has Pro at $42.50/month and Max at $272/month billed annually, plus pay-as-you-go credits for agent runs
  • Includes a Docker sandbox. Listed for 48 of 194 tools in this category.
  • Supports headless CI workflows. Listed for 60 of 194 tools in this category.

“Forty-five integrations, from the project that once promised to replace the person who would have set them up.”

Website Docs 188k starsCompare vs…Dispute a fact
Appeal a claim or request ownership transfer

What it is

AutoGPT is a platform for building, deploying and running AI agents that execute complete workflows. You describe the outcome in plain English or use the visual Build editor, then run agents on demand, on a schedule or from triggers, with a Marketplace of community agents. The original standalone agent lives on as AutoGPT Classic in the same repository.

Specification

Source verification

Row snapshot checked 2026-09-03. Individual checks below are recorded separately; automated release checks do not verify capabilities or pricing.

readme
Needs individual review
pricing
Needs individual review
install
Needs individual review
protocols
Needs individual review
models
Needs individual review

Architecture

Type
Agent harness
Runsunsourced
local, cloud
Platforms
macos, linux, windows, web
Context windowsrc ↗
not documented
Languages
any

Models

Backbonesrc ↗
OpenAI, Claude, Llama, hundreds of models via integrations
Bring your own model
Yes
Local models
Yes

Protocols

MCP clientsrc ↗
Yes
MCP server
No
OpenAPI tools
No

Capabilities

Terminal commandsunsourced
No
Multi-file edits
No
Git operations
No
Browser control
No
Sandboxed execution
Yes
Multi-agent
Yes
Headless / CI
Yes

Cost

Modelsrc ↗
mixed
Starts at
$42.5/mo
Free tier
Yes
Bring your own key
Yes

Self-hosting is free with your own API keys; the managed platform has Pro at $42.50/month and Max at $272/month billed annually, plus pay-as-you-go credits for agent runs

Openness

Open sourceunsourced
Yes
License
Polyform Shield 1.0 (platform) / MIT (classic and other components)
First release
unknown
open-sourceplatformvisual-buildermarketplacescheduledtriggerscloudself-hosted

Los Agentes on AutoGPT

Who are they?
The ruling
El JuezThe judge

The panel lands inside 1.25 points and nobody is enthusiastic; El Crítico names why, the 187,000 stars belong to a 2023 agent now sitting in a classic folder.

Trial only
Reasoning and trade-offs · AI analysis

Agreement without warmth: 1.25 points from El Hacker's 6.00 to La Jefa's 4.75. El Crítico states the mismatch, the stars belong to a 2023 agent that now lives in a classic folder while the product on the landing page is a workflow builder with no shell. El Amigo puts it plainly, this one never touches a file.

El Amigo's boundary is the ruling: judged as a workflow platform it is a fair product, judged as the autonomous agent the stars remember it is gone. La Jefa blocks it and she is right to: the Team tier is marked coming soon. Trial only, the exit criterion being that tier shipping.

Agree with El Juez?
El AmigoThe friend

Pick AutoGPT if the job is a scheduled workflow across Gmail, Slack and Jira built in a block editor; pick goose or Claude Code if the job is code, because this one never touches a file.

5.8
Reasoning and trade-offs · AI analysis

You will like this if what you want is automation, not engineering: describe an outcome, or drag blocks together in the Build editor, then run it on a schedule or from a trigger and check the run dashboard in the morning. The daily trait is that it behaves like an automation product, wired to the SaaS tools your company already pays for.

You will not like it if you came for the original agent that wrote and ran code. Today's platform has no terminal, no file edits and no repository awareness. Pick it for ops workflows. Pick goose or Claude Code when the task lives in a repo.

reliability
6
usefulness
6
cost
5
longevity
6
Agree with El Amigo?
El CríticoThe critic

187,000 stars were earned by a 2023 agent that now lives in a classic folder, while the product on the landing page is a workflow builder with no shell and MCP reduced to a block.

5.3
Reasoning and trade-offs · AI analysis

The risk is a borrowed reputation. The 187k stars and 46k forks belong to the original autonomous agent, which the repository now keeps in a classic folder. The platform sold on the landing page is a different product: a block-based workflow builder with no terminal, no file edits, and MCP only as a block you wire yourself. Evaluate it as a new product with a short track record, because that is what it is.

The consequence: do not let the star count stand in for a reference call. What it does right is isolation. Agent runs execute in Docker, so a bad block does not run on your laptop.

reliability
5
usefulness
5
cost
5
longevity
6
Agree with El Crítico?
El ProfesorThe professor

Four documented surfaces, AutoPilot, Agents, Build and Marketplace, and one benchmark tool, agbenchmark, which ships with the original agent's code rather than with the platform being sold.

5.5
Reasoning and trade-offs · AI analysis

The documented architecture has four surfaces. 1. AutoPilot turns a natural-language outcome into an agent. 2. Build is a visual editor where blocks are wired together, so the plan is a graph the user drew rather than one the model inferred. 3. Agents is a dashboard of runs, costs and statuses. 4. Marketplace distributes community agents.

The verification story is thinner. The only benchmark tooling, agbenchmark, sits with the original agent's code, and no published number describes how the platform's agents perform. The design is principled where the user draws the graph; where AutoPilot draws it, the documentation does not say how the result is checked.

reliability
6
usefulness
5
cost
5
longevity
6
Agree with El Profesor?
La InversoraThe investor

Significant Gravitas pivoted from the agent that started the category to a workflow platform priced at $42.50 and $272 a month, billed annually, with credits on top.

5.8
Reasoning and trade-offs · AI analysis

This is the rare pivot that kept the brand and swapped the product. The agent that started the category became a workflow platform with Pro at $42.50 a month and Max at $272, both billed annually, plus pay-as-you-go credits for runs. Anchoring at $272 for individuals is a pricing-power test, and annual billing says they want the cash up front.

The moat is the name and the marketplace, neither of which stops Zapier or Make from adding an agent block. Likely acquirer: an automation incumbent buying the brand. Likely pivot: enterprise plans sold through a sales contact. Position: cautious, wait for a revenue signal.

reliability
6
usefulness
6
cost
5
longevity
6
Agree with La Inversora?
La JefaThe CTO

Sixty seats need the Team tier with admin roles and centralized billing, and that tier is marked coming soon, so today it is sixty individual subscriptions and a credit wallet.

4.8
Reasoning and trade-offs · AI analysis

The demo is a chat that builds an automation and runs it on a schedule. Procurement finds that the Team tier, the one with multi-user workspaces, admin roles and seat management, is listed as coming soon. Today that means sixty individual plans, sixty credit wallets metered per run, and an SSO story that exists on the cloud platform as OAuth 2.0 but without SCIM.

CI fit is nil, because this is not a code tool, and the onboarding burden is low for the same reason. Support is email on the standard tier. Not yet. Revisit when the Team tier has a price.

reliability
5
usefulness
5
cost
4
longevity
5
Agree with La Jefa?
El HackerThe tinkerer

Polyform Shield on the platform folder forbids selling it as a competing hosted service, the classic agent stays MIT, MCP is a canvas block rather than a file I version, and self-hosting with local Llama works.

6.0
Reasoning and trade-offs · AI analysis

Read the license before the README. The platform directory is Polyform Shield 1.0, which lets me run it for my own use but not sell it as a competing hosted service, and the rest is MIT. That is honest source-available, not open source, and I respect the honesty more than the marketing.

Self-hosting is a shell script and my own keys, and the model list includes Llama, so a local endpoint is in play. MCP is a block on a canvas, not a file I can version, which is backwards. I can fork it for myself; I cannot fork it into a business. That is the deal.

reliability
6
usefulness
5
cost
7
longevity
6
Agree with El Hacker?