Model choice · Browser Use Agents

Pick the model.
Keep the browser.

One agent harness. A choice of models. Find the right balance for the work you actually run.

Managed agent models

Per 1M tokens · USD

Managed models and input and output token prices
ModelInputOutput
Claude Opus 5$6.00$30.00
Claude Sonnet 5$2.40$12.00
Grok 4.5$2.40$7.20
GPT-6 Astra$12.00$60.00
GPT-5.6$6.00$36.00
GPT-5.6 LunaDEFAULT$0.24$1.44
Gemini 3.6 Flash$1.80$9.00
MiniMax M3$0.36$1.44

Choose per run

Start with GPT-5.6 Luna, the default. Change the model as your task's accuracy, latency, and cost requirements change.

Model documentation

Use your own key

Eligible plans support your model-provider key. Provider token costs, Browser Use orchestration, browser time, and traffic remain separate charges.

Understand the charges

Purpose-built models

Browser Use also develops models trained for browser work. BU 2.0 has its own model API rates, separate from managed V4 runs.

Purpose-built model pricing

Choose with evidence

Compare on browser work.

Accuracy and cost depend on the task, model, and harness. Start with a matched benchmark, then test your own workflow.

Browser Use: 82% of tasks solved at 17¢ each. That is 20 points better than Opus 5, which costs 20× more per solved task.

Toolstrict accuracyCost
Browser Use82%$0.17
Opus 562%$3.40
Gemini 3.1 Pro59%$2.20
Sonnet 559%$1.55
GPT-5.652%$1.10
Gemini 3.6 Flash46%$0.62
GPT-537%$0.44
Internal Bench Hard · updated 2026-08-01 · all benchmarks

Earlier Browser Use 1.0 latency measurements use the OnlineMind2Web workload. Read that experiment for its model versions and methodology.