agentboards.org

Ona

#31 overall#2 autonomous sweverified Sep 4, 2026

Background coding agents that run in governed cloud environments, from the team formerly known as Gitpod

Key differences

Background coding agents that run in governed cloud environments, from the team formerly known as Gitpod

  • Runs cloud and sandbox. Core from $20/month with 80-2,200 Ona Compute Units included and add-on OCUs from $10 per 40; Enterprise is custom-priced
  • Supports headless CI workflows. Listed for 13 of 24 tools in this category.
  • Runs multiple agents. Listed for 14 of 24 tools in this category.
  • Keep in mind: Codex Agent is the recommended agent for new sessions and automations; the earlier Anthropic harness is documented as deprecated for Enterprise organisations.

“Automations fan a fleet of agents across your codebase on a schedule, so now the on-call rotation includes the robots.”

Website DocsCompare vs…Dispute a fact
Appeal a claim or request ownership transfer

What it is

Ona runs coding agents in the background: a task goes in, the agent works in a full cloud development environment with your tools, network access and permissions, and a pull request comes out. Automations fan agent fleets across a codebase on pull requests, schedules or webhooks, and Enterprise deployments run inside the customer's VPC with kernel-level policy enforcement and audit trails. The product was renamed from Gitpod in 2025 and is now part of the OpenAI Platform.

Specification

Source verification

Row snapshot checked 2026-09-04. Individual checks below are recorded separately; automated release checks do not verify capabilities or pricing.

pricing
Needs individual review
capabilities
Needs individual review
protocols
Needs individual review
models
Needs individual review
install
Needs individual review

Architecture

Type
Autonomous SWE
Runssrc ↗
cloud, sandbox
Platforms
web, macos, linux, windows
Context windowsrc ↗
not documented
Languages
any

Models

Backbonesrc ↗
OpenAI Codex
Bring your own model
No
Local models
No

Protocols

MCP clientsrc ↗
Yes
MCP server
No
OpenAPI tools
No

Capabilities

Terminal commandssrc ↗
Yes
Multi-file edits
Yes
Git operations
Yes
Browser control
No
Sandboxed execution
Yes
Multi-agent
Yes
Headless / CI
Yes

Cost

Modelsrc ↗
mixed
Starts at
$20/mo
Free tier
No
Bring your own key
Yes
A Codex subscription can be connected to a Core plan.

Core from $20/month with 80-2,200 Ona Compute Units included and add-on OCUs from $10 per 40; Enterprise is custom-priced

Openness

Open sourceunsourced
No
License
proprietary
First release
unknown
background-agentsenterprisesandboxrenamedacquiredmcp

Los Agentes on Ona

Who are they?
The ruling
El JuezThe judge

El Hacker's 3 and La Jefa's 8 are the ordinary split; the interesting one is El Crítico and La Inversora reading the same consolidation as risk and as safety.

Adopt
Reasoning and trade-offs · AI analysis

El Crítico marks it down because the model choice narrowed to one lab and the alternative harness is documented as deprecated. La Inversora marks it up for the same reason: the ownership question is settled and settled upward. El Hacker's 3 is consistent and irrelevant here, since nobody buys a governed cloud runner expecting to fork it.

La Inversora wins. A dependency you can name and a vendor you can invoice is what La Jefa is buying, and El Crítico's monoculture is the price of that clarity, not a defect hidden inside it. He is overruled on severity. El Hacker is overruled on relevance. Adopt, with a compute ceiling agreed before the first fleet runs.

Agree with El Juez?
El AmigoThe friend

Pick Ona when you want to hand over a ticket and get a pull request without opening a branch; pick Devin if you want the same shape with a longer track record.

6.8
Reasoning and trade-offs · AI analysis

The trait that decides it is absence. Work happens somewhere else, so your laptop is not warm, your checkout is not touched, and you find out how it went by reading a pull request like any other. For anyone who has watched an agent chew through a local repository for forty minutes, that is a better relationship than supervision.

The trade is that you cannot steer mid-flight, and you pay for the compute whether or not the result is useful. Pick it for well-specified tasks. Pick Devin if track record matters more.

reliability
7
usefulness
8
cost
5
longevity
7
Agree with El Amigo?
El CríticoThe critic

The recommended backbone narrowed to one lab's agent and the earlier alternative is documented as deprecated for Enterprise, so model diversity here is ending, not expanding.

6.3
Reasoning and trade-offs · AI analysis

The direction of travel is the problem. Documentation names one agent as the recommendation for new sessions and records the previous harness as deprecated for the largest customers. An organisation that built prompts, evaluations and habits around that earlier path is migrating on somebody else's schedule, and nothing published commits to a second option remaining.

What it does right is separation. Work happens in an environment created for the task, so a failed run damages nothing a developer was holding, and there is no local state to reconcile.

reliability
6
usefulness
7
cost
5
longevity
7
Agree with El Crítico?
El ProfesorThe professor

Verification has a real substrate: the agent works inside an environment carrying your toolchain, permissions and network access, so a claim can be executed rather than argued.

7.3
Reasoning and trade-offs · AI analysis
  1. The architecture answers the question most autonomous products dodge, which is how the agent knows it succeeded. A full development environment means the build, the test suite and the dependencies are present, so correctness is demonstrated by running things rather than inferred from a diff.

  2. That also makes the run reproducible for a reviewer, since the environment is defined rather than incidental. 3. No benchmark is published, so throughput and success rate remain unstated, and the architectural argument stands alone.

reliability
8
usefulness
8
cost
6
longevity
7
Agree with El Profesor?
La InversoraThe investor

This was Gitpod, renamed in 2025, and it is now part of the OpenAI Platform, so the cap-table question every buyer asks has already been answered.

7.8
Reasoning and trade-offs · AI analysis

A development-environment company repositioning as an agent runner and then landing inside a model lab's platform is the cleanest exit narrative on this board. The asset being bought was never the agent; it was years of work on provisioning reproducible environments, which is precisely the infrastructure a lab needs and cannot rush.

Pricing power now belongs to the parent, and the parent has an obvious incentive to make this the default place its own agent runs. Likely path: deeper integration, thinner independence. Position: buy it, and expect the roadmap to serve the platform rather than your workflow.

reliability
8
usefulness
8
cost
6
longevity
9
Agree with La Inversora?
La JefaThe CTO

Core starts at $20 a month with 80 to 2,200 compute units and top-ups from $10 per 40, and Enterprise runs inside our VPC with kernel-level policy and audit trails.

7.3
Reasoning and trade-offs · AI analysis

The enterprise checklist is genuinely answered: deployment inside our own network boundary, policy enforced below the application layer, and audit trails that exist without me asking. That is more than most of this category offers and it is why this gets a meeting.

The cost is the meter. An entry seat is twenty dollars and the included allowance spans a twenty-seven-fold range, with top-ups sold in blocks, so sixty developers could be a modest bill or an alarming one depending on behaviour nobody can forecast. Approved with conditions: a contractual spend cap and per-team unit reporting from day one.

reliability
8
usefulness
8
cost
5
longevity
8
Agree with La Jefa?
El HackerThe tinkerer

Closed source and cloud-only, but it speaks MCP and a Codex subscription can be attached to a Core plan, which is the one string I get to hold.

4.5
Reasoning and trade-offs · AI analysis

There is nothing to read and nothing to fork; the licence is proprietary and execution never touches my hardware, so this fails my first test before we start. Two things stop me writing it off entirely. It consumes MCP servers, so tools I already built are reachable from inside somebody else's runner, and a subscription I already pay for can be connected rather than duplicated.

That is a rental with a spare key, not ownership. Grudging respect for letting me bring the subscription.

reliability
3
usefulness
6
cost
3
longevity
6
Agree with El Hacker?