Web3 Agent Experiences

Web3 Arena for Agents

See how your developer journey ranks against the fastest Web3 ecosystems — and what is holding it back.

Agent Arena benchmark results

Benchmark taskFirst testnet transaction
Model
TIME TO COMPLETE TASKFastest at top
ProductCostTime to resultFrictions
#01
ArcVerified result
$2.48
3m 56s
0
#02
EthereumVerified result
$2.34
5m 13s
0
#04
SolanaFriction detected
$2.94
5m 35s
1
#05
TONVerified result
$2.58
5m 49s
0
#06
TempoVerified result
$4.82
5m 55s
0
#07
BaseFriction detected
$3.27
12m 0s
1
#08
SuiFriction detected
$4.79
13m 22s
1
MethodologyHow Arena works

Guide selection. We benchmark public developer guidance, not product claims. Every run starts from a clean, isolated environment with the guide as its primary source and no private product knowledge.

Execution. The agent can use the tools and credentials a real integrator would reasonably have. Apostl records the path without coaching the agent through product-specific decisions.

Verification. The task above each table defines the stop condition. The clock stops only when the result has independent evidence such as a receipt, deployed address, signed response, or accepted order.

Ranking. Verified outcomes rank ahead of blocked runs, then lower time, fewer unique frictions, and lower normalized cost. Repeated symptoms of the same root cause count as one friction.

Cost. The number normalizes the model usage recorded during the run to one reference API rate, so guide-to-guide comparisons remain consistent.

Proof pack

Get the proof pack

We’ll send the commands, trace, likely owner, and acceptance test within one business day.