← All leaderboards
Agentic browsing
What this board tests: a natural-language goal handed to an autonomous browser agent. Every provider below ran the identical corpus on 2026-06-22. The composite blends measured quality, latency p50, cost per successful call and error rate. Method and corpora: how we measure.
| Rank | Provider | Composite | Task success | Cost | Latency p50 | Field note |
|---|---|---|---|---|---|---|
| 1 | 7.99/10 | 1.00 | $0.05 / call | 47.3 s | Front-runner in autonomous NL-agent mode. | |
| 2 | 5.66/10 | 0.67 | $0.05 / call | 70.2 s |
Low scores are published alongside high ones — a ranking you can only trust if nothing gets hidden. The gateway's fee sits on top-ups, not per call, so the order above earns us nothing either way.