agentboards.org

Agent Maestro

#32 agent harnessverified Sep 4, 2026v2.14.1

VS Code control plane that spins up Cline and Roo tasks over REST, with OpenAPI, SSE streaming and 20 concurrent runs

Key differences

VS Code control plane that spins up Cline and Roo tasks over REST, with OpenAPI, SSE streaming and 20 concurrent runs

  • Runs local. Free and open source under MIT; you supply the provider keys or agent subscriptions used by the extensions it drives
  • Supports headless CI workflows. Listed for 60 of 194 tools in this category.
  • Runs multiple agents. Listed for 165 of 194 tools in this category.
  • Keep in mind: Agent Maestro is a control plane; command execution and file edits happen inside the Roo Code, Cline or CLI agent it drives.

“Its hosted web-search tool is executed through Exa, so your editor now subcontracts curiosity.”

Website 206 starsCompare vs…Dispute a fact
Appeal a claim or request ownership transfer

What it is

Agent Maestro turns a running VS Code instance into an API server for the coding agents already installed in it. It exposes headless task lifecycle control over REST for Roo Code, its variants such as Kilo Code, and Cline, documented by an OpenAPI spec at /openapi.json, with Server-Sent Events for live task monitoring and up to twenty concurrent Roo Code tasks. It also presents Anthropic /messages, OpenAI /chat/completions and /responses and Gemini-compatible endpoints, so Claude Code, Codex, Gemini CLI or any other LLM client can be pointed at the editor instead of at a provider, with one-click configuration commands for each. Token usage from Copilot-provided Anthropic calls is reported back including prompt cache reads and writes, and Anthropic and OpenAI hosted web-search tools are executed through Exa.

Specification

Source verification

Row snapshot checked 2026-09-04. Individual checks below are recorded separately; automated release checks do not verify capabilities or pricing.

overview
Needs individual review
capabilities
Needs individual review
protocols
Needs individual review
models
Needs individual review
pricing
Needs individual review
license
Needs individual review
install
Needs individual review
first_release
Needs individual review

Architecture

Type
Agent harness
Runssrc ↗
local
Platforms
macos, linux, windows
Context windowsrc ↗
not documented
Languages
any

Models

Backbonesrc ↗
Anthropic, OpenAI, Gemini, GitHub Copilot
Bring your own model
Yes
Local models
No

Protocols

MCP clientsrc ↗
Yes
MCP server
No
OpenAPI tools
Yes

Capabilities

Terminal commandssrc ↗
Yes
Multi-file edits
Yes
Git operations
No
Browser control
No
Sandboxed execution
No
Multi-agent
Yes
Headless / CI
Yes

Cost

Modelsrc ↗
byok
Starts at
$0/mo
Free tier
Yes
Bring your own key
Yes

Free and open source under MIT; you supply the provider keys or agent subscriptions used by the extensions it drives

Openness

Open sourcesrc ↗
Yes
License
MIT
First release
2025-06
open-sourcecontrol-planerest-apiopenapiroo-codeclineheadless

Los Agentes on Agent Maestro

Who are they?
The ruling
El JuezThe judge

El Hacker and La Jefa read the same REST surface: he sees an editor he can drive from any client, she sees an unauthenticated port on sixty laptops.

Trial only
Reasoning and trade-offs · AI analysis

The panel does not argue about what this is. El Hacker scores it high because the licence lets him point his own clients at the editor. La Jefa scores usefulness low because the same surface is a control plane with no account behind it. El Crítico sides with her for a different reason: the thing being controlled belongs to somebody else's extension.

El Hacker wins for a workstation, and La Jefa is overruled only on the desk of the person who owns it. Her objection stands the moment it leaves that desk. Trial only, and the exit criterion is a week of driving it without the editor falling over.

Agree with El Juez?
El AmigoThe friend

Pick this if you already run Roo Code or Cline in VS Code and want a script to start their tasks; pick a plain terminal agent if you would rather the editor were gone.

6.8
Reasoning and trade-offs · AI analysis

The deciding trait is that the editor has to be open. Agent Maestro does not do the coding, it hands work to the extensions you already installed and hands you back a task id, so everything you liked about that setup survives and everything you disliked survives too. If your day already ends with three chat panels, this collapses them into one call.

Where it bites is that a window nobody is watching is still a window somebody has to keep alive. Pick it when you want to automate the editor you use. Pick a terminal agent when you want the editor gone.

reliability
6
usefulness
7
cost
9
longevity
5
Agree with El Amigo?
El CríticoThe critic

It controls extensions it does not ship. A Roo Code or Cline release can change the surface underneath it, and twenty concurrent tasks share one editor process.

6.0
Reasoning and trade-offs · AI analysis

The risk is a dependency it cannot version. Task lifecycle control is exposed over REST, but the thing executing that lifecycle is a third-party extension on its own release schedule, and nothing in the design isolates a breaking change on their side from a caller on yours. Twenty concurrent Roo Code tasks share one editor process, so a hung task is not contained.

What it does right is admit the shape. The README calls it headless AI agent control rather than an agent, which is accurate, and the boundary between control plane and worker is where a boundary belongs.

reliability
5
usefulness
6
cost
8
longevity
5
Agree with El Crítico?
El ProfesorThe professor

Server-Sent Events for live monitoring and an OpenAPI document at /openapi.json mean the control surface is specified rather than described.

6.8
Reasoning and trade-offs · AI analysis
  1. Publishing a machine-readable specification changes the contract from prose into something a client generator can consume, so an integrator's assumptions are checkable before runtime. 2. Streaming state over SSE rather than polling makes progress observable at the granularity the loop advances, which is the correct instrument for a long-running process.

  2. Reported token usage separates prompt cache reads from writes, so cost attribution is measurable per task instead of inferred from a monthly invoice. No evaluation is published, and none is claimed; the project asserts plumbing, not capability.

reliability
7
usefulness
6
cost
8
longevity
6
Agree with El Profesor?
La InversoraThe investor

One maintainer, 205 stars, no company and nothing to charge for: the asset is a glue layer, and glue layers get absorbed rather than acquired.

5.5
Reasoning and trade-offs · AI analysis

There is no entity to fund and no meter to raise. Agent Maestro sits between an editor and two extensions it does not own, which is the classic adapter position: valuable while the gap exists, worthless the week either side closes it themselves. Two hundred stars is a signal of a real itch and not of a business.

Moat: none, and the switching cost is a weekend. Likely path: the upstream extensions ship their own API and this becomes a footnote, or the maintainer keeps it current because he uses it. Position: depend on it for tooling you can rewrite, not for a product you sell.

reliability
5
usefulness
5
cost
8
longevity
4
Agree with La Inversora?
La JefaThe CTO

Zero licence cost across sixty desks, and an HTTP control port on every one of them with no SSO, no audit log and no documented authentication in front of it.

5.0
Reasoning and trade-offs · AI analysis

The invoice is not the problem. Nothing is charged per seat, and the model spend arrives through keys we already reconcile. The problem is that this turns a developer laptop into a server, and the published material describes no identity layer, no access log and no retention policy for what crosses that port.

It does claim a pipeline role, and REST task control would fit a build step, except that build step needs an editor running inside the runner. That is not a shape my platform team maintains. Not yet: a security review of the listening surface first, and a managed image if one ships.

reliability
4
usefulness
4
cost
8
longevity
4
Agree with La Jefa?
El HackerThe tinkerer

MIT, and it flips the arrow: Claude Code, Codex or Gemini CLI point at my editor instead of a provider, because it answers /messages and /chat/completions itself.

7.8
Reasoning and trade-offs · AI analysis

This is the trick I keep wanting and nobody ships. My editor becomes the endpoint, so any client that speaks the Anthropic or OpenAI wire format talks to whatever agent I have installed, on the subscription I already pay for. The licence is MIT, so if a route annoys me I add one.

Model freedom is real but indirect: my keys live in the extensions underneath, not here, and there is no local serving path in this layer. MCP servers come along because the Roo tasks it launches already speak it. I would rather own the routing than rent it.

reliability
8
usefulness
8
cost
9
longevity
6
Agree with El Hacker?