> ## Documentation Index
> Fetch the complete documentation index at: https://browseruse-0aece648-codex-docs-supported-exports.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

> ## Agent Instructions
> Use https://docs.browser-use.com/llms.txt and its linked .md pages for current documentation. The managed full bundle is https://docs.browser-use.com/.well-known/llms-full.txt and can be cached for up to 24 hours. Do not use the obsolete /cloud/llms*.txt or /open-source/llms*.txt static exports.
> Choose Cloud API V4 for new agent integrations; V2 is the lower-cost option for simple tasks. Keep V3 examples explicitly versioned. The open-source browser-use library and hosted browser-use-sdk have different APIs.
> Cloud authentication uses X-Browser-Use-API-Key, without a Bearer prefix. Install or upgrade browser-use-sdk and use its explicit v4 import for V4. Check the published OpenAPI reference for request fields; do not invent SDK support for new fields.
> Cloud concurrency and HTTP request rate are separate. Read GET /api/v2/billing/account for the key’s projectId, concurrentSessionLimit, activeSessionCount, and credit balance, including when using V4. Keys in one project share capacity and credits; rateLimit is a legacy concurrency alias, not requests per second.
> Keep the highest applicable existing, legacy-plan, and spend-tier concurrency grant. Current spend tiers are 10 / 50 / 250 / 500 / 1000 at $0 / $200 / $1000 / $5000 / $25000 in qualifying project payments. Legacy or externally billed projects can follow different billing paths; trust the account limit. See https://docs.browser-use.com/cloud/guides/concurrency.md.
> Budget polling across the project: the standard general bucket is 25 requests/second, including V4 event reads and full run reads. Selected status reads have a separate higher bucket. Use bounded workers, stagger polls, respect Retry-After, and drain hasMore event pages after terminal status. A busy V4 session returns 409; its queue holds 10 pending messages and is not a project-wide batch queue.
> A completed run or closed CDP connection does not immediately stop its cloud browser. Stop unneeded owned browsers with PATCH /api/v4/browsers/{id} and {"action":"stop"}. A client wait timeout does not cancel the server-side run.
> Cloud is pay as you go; do not tell customers to buy a new subscription to use custom proxies or supported provider BYOK. Usage funding and model eligibility still apply. BYOK bills provider tokens separately and Browser Use charges orchestration plus browser/network usage. See https://docs.browser-use.com/cloud/guides/billing.md.
> Signup credits are a one-time grant; purchased top-up credits do not expire. Check the API key’s project before diagnosing missing credits. API-key monthly spending caps are soft limits, not a strict prepaid wallet; concurrent or already-running work can exceed them. Auto recharge has separate trigger and purchase amounts and can charge immediately when enabled below the threshold. Use https://browser-use.com/pricing for current rates.

# Choosing an Agent

> Compare Browser Use Cloud agents on accuracy, speed and price, and pick the right one for your task.

Browser Use Cloud offers three agents that can complete your tasks. Use V4 when accuracy matters most, V3 when speed and cost matter most, and V2 only for older integrations.

## At a glance

|                        | V4 Agent                 | V3 Agent                              | V2 Agent                            |
| ---------------------- | ------------------------ | ------------------------------------- | ----------------------------------- |
|                        | Most accurate            | Fastest                               | Legacy                              |
| Verdict                | Choose for complex tasks | Choose for quick, cost-optimized work | Keep only for existing integrations |
| Accuracy on hard tasks | **76%**                  | 67%                                   | 54%                                 |
| Speed                  | slower, more thorough    | **fastest on identical tasks**        | quick on small tasks                |
| Cost level             | \$\$\$                   | \$\$                                  | \$                                  |

### V4 Agent — most accurate

V4 writes and runs code to complete long, multi-step browser tasks. It can research across many pages, compare options, follow detailed instructions, collect records, and save the results as a spreadsheet or in any format.

**Best for:**

* Bulk data collection
* Complex tasks that span many pages and sites
* Long, difficult instructions

**Example tasks:**

* "Collect every document with its title, reference, and deadline."
* "Create a spreadsheet of all the pricing plans across these five competitors."

The most thorough option, but it takes its time.

### V3 Agent — fastest

V3 looks at the page and acts step-by-step like a human would. Use only when cost and speed matter more than accuracy.

**Best for:**

* Simple tasks, such as finding publicly available information
* Workloads where cost and speed are essential
* Open-ended research

**Example tasks:**

* "Get a shipping quote for a 3 lb package between two zip codes."
* "Check whether this product is in stock and what it costs right now."

Faster but less reliable.

### V2 Agent — legacy

Our first-generation agent remains available for compatibility. It is based on our open-source repository, but it is not actively maintained and is less accurate.

Starting something new? Use V4 instead; it's far more accurate.

## V4 completes the hardest tasks

V4 completed 76% of difficult, real-world browser tasks — nine points ahead of V3 and twenty-two ahead of V2.

<img src="https://mintcdn.com/browseruse-0aece648-codex-docs-supported-exports/pg6YX8jO9xwCWYaf/cloud/images/agent-accuracy.png?fit=max&auto=format&n=pg6YX8jO9xwCWYaf&q=85&s=c2e7ca0c2565baea61a3b95ff9b6099c" alt="Browser Use Cloud success rate on hard web tasks: V2 Agent 54.25%, V3 Agent 67.25%, V4 Agent 76.47%" width="2400" height="1433" data-path="cloud/images/agent-accuracy.png" />

## Typical tasks cost \$0.70–\$1.64

Measured with the same model — Claude Opus 4.7 — across the last 30 days of production usage, the typical customer's successful task costs \$0.70 on V2, \$0.87 on V3, and \$1.64 on V4, taking about 1, 2, and 5 minutes respectively. V4 costs more per task because customers use it for more complex work, and V2's tasks run shortest only because it gets the simplest ones. On identical hard tasks, V3 is the fastest agent (about 6 minutes, versus about 10 for V2 and V4).

|               | V2 Agent | V3 Agent | V4 Agent |
| ------------- | -------- | -------- | -------- |
| Cost per task | \$0.70   | \$0.87   | \$1.64   |
| Time per task | \~1 min  | \~2 min  | \~5 min  |

<img src="https://mintcdn.com/browseruse-0aece648-codex-docs-supported-exports/pg6YX8jO9xwCWYaf/cloud/images/agent-real-cost.png?fit=max&auto=format&n=pg6YX8jO9xwCWYaf&q=85&s=76da7229d62838a761e91a849571c8ac" alt="What a task costs with Claude Opus 4.7 from real usage: V2 $0.70, V3 $0.87, V4 $1.64 per successful task" width="2400" height="1433" data-path="cloud/images/agent-real-cost.png" />

## Comparing accuracy, speed, and cost

<img src="https://mintcdn.com/browseruse-0aece648-codex-docs-supported-exports/pg6YX8jO9xwCWYaf/cloud/images/agent-comparison.png?fit=max&auto=format&n=pg6YX8jO9xwCWYaf&q=85&s=9e85028ca08e7afc5a9b78422e2ad06a" alt="Comparing agent accuracy, speed, and cost: V4 is most accurate at $1.64 per task, V3 is fastest at $0.87, V2 is cheapest at $0.70 but least accurate" width="2400" height="1549" data-path="cloud/images/agent-comparison.png" />

## How we measured this

In July 2026, we ran each agent four times on Internal Bench Hard: 106 difficult tasks on real websites. Every run used Claude Opus 4.8, and the same independent judge scored each result. We left out tasks blocked by websites' anti-bot measures. The speed and cost figures are averages; simpler tasks usually run faster and cost less. Real-world cost and speed figures are the median customer's successful task with Claude Opus 4.7 over the last 30 days of production, our own testing excluded.
