Skip to content
GitHub

Product

Benchmarks Partners

Resources

Docs Blog

The default benchmark platform for infrastructure

Independent benchmarks for compute

The best way to evaluate your compute provider performance.

Median TTI (burst) over time

Sandboxes

Live
Declaw140ms
Arker177ms
Northflank252ms
Beam290ms
Isorun303ms
Median TTI (burst) · last 8 runsFull history →
Backed by
LatitudeGoogle Cloud RunBrowserbaseTigrisNeonGitbookNamespace
Methodology

How the Numbers Are Made

Every benchmark is reproducible: open code, fixed workloads, results published in full

1

One workload, every provider

The same script, payload, and model hit every provider, addressed the way its own API expects. Nobody gets a tuned path.

2

Repeated on a schedule

Every suite runs many iterations per provider, automated in GitHub Actions on a recurring schedule, so results reflect steady state rather than one lucky request.

3

Scored on the tail, not the best case

Each composite score blends median and tail latency against a fixed ceiling, then multiplies by success rate — providers that fail often lose real points.

4

Raw results published

Every run is committed as JSON to the public benchmarks repo, next to the code that produced it. Nothing is cherry-picked or held back.

Trusted by the best

Independent benchmarks of every major sandbox provider, sponsored by industry leaders.

LatitudeGoogle Cloud RunBrowserbaseTigrisNeonGitbookNamespace

Start Benchmarking

We provide independent benchmarking for all compute providers.