Skip to content

Claude Opus 5 provider comparison

anthropic/claude-opus-5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

Endpoints:
6
Context:
1M
Input:
text, image, file
Output:
text
Compare this model in a set →

At a glance

Comparison winners

Based on 1,000 input and 500 output tokens using the latest published median speed data.

Best cost/speed trade-off

Azureazure/us-east-2

$11.6667 / 1M tokens

9.55 s estimated response

Cheapest

Amazon Bedrockamazon-bedrock/claude-on-aws

$11.6667 / 1M tokens

$0.0175 for the sample request

Fastest

Azureazure/us-east-2

9.55 s estimated response

$11.6667 / 1M tokens

Visual comparison

Price and performance charts

Compare published endpoint pricing and estimated response time visually.

Effective price by endpoint

USD per 1 million tokens for input and output. Lower is better.

Amazon Bedrock has the lowest estimated cost for 1,000 input and 500 output tokens at $0.0175.

Estimated cost vs. response time

Based on 1,000 input and 500 output tokens. Cost is shown in USD per 1 million tokens. Lower and further left is better; the single best cost/speed trade-off appears at full opacity.

5 providers have complete pricing and speed data. Azure is closest to the ideal combination of lowest cost and fastest response.

Server-rendered comparison

Provider endpoints

Prompt and completion prices are effective USD per token as published by OpenRouter. Missing speed data never removes an endpoint.

Endpoint comparison for Claude Opus 5
Provider / routePricing ↗BenchmarksAPI support
Provider / route
amazon-bedrock/claude-on-aws
1M contextunknown128K max output
Cheapest
Pricing
Input
$5 / 1M
Output
$25 / 1M
Cache
$0.5 / 1M
Blended— Best comparable value
$11.6667 / 1M
Benchmarks
Speed— Worst comparable value
13.41 s
TTFT
3.60 s
TPS— Worst comparable value
51.0 tok/s
Uptime
99.99%
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls
10 parameters

include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, stop, structured_outputs, tool_choice, tools, verbosity

Provider / route
anthropic
1M contextunknown128K max output
Pricing
Input
$5 / 1M
Output
$25 / 1M
Cache
$0.5 / 1M
Blended— Best comparable value
$11.6667 / 1M
Benchmarks
Speed
11.35 s
TTFT— Worst comparable value
3.77 s
TPS
66.0 tok/s
Uptime— Best comparable value
100.00%
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls
10 parameters

include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, stop, structured_outputs, tool_choice, tools, verbosity

Provider / route
azure/us-east-2
1M contextunknown128K max output
FastestBest cost/speed trade-off
Pricing
Input
$5 / 1M
Output
$25 / 1M
Cache
$0.5 / 1M
Blended— Best comparable value
$11.6667 / 1M
Benchmarks
Speed— Best comparable value
9.55 s
TTFT— Best comparable value
2.15 s
TPS— Best comparable value
67.5 tok/s
Uptime
N/a
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls
12 parameters

include_reasoning, max_completion_tokens, max_tokens, reasoning, reasoning_effort, response_format, stop, structured_outputs, temperature, tool_choice, tools, verbosity

Provider / route
amazon-bedrock
1M contextunknown128K max output
Pricing
Input
$5 / 1M
Output
$25 / 1M
Cache
$0.5 / 1M
Blended— Best comparable value
$11.6667 / 1M
Benchmarks
Speed
10.97 s
TTFT
2.35 s
TPS
58.0 tok/s
Uptime— Worst comparable value
99.97%
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls
10 parameters

include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, stop, structured_outputs, tool_choice, tools, verbosity

Provider / route
google-vertex/global
1M contextunknown128K max output
Pricing
Input
$5 / 1M
Output
$25 / 1M
Cache
$0.5 / 1M
Blended— Best comparable value
$11.6667 / 1M
Benchmarks
Speed
11.39 s
TTFT
2.30 s
TPS
55.0 tok/s
Uptime— Best comparable value
100.00%
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls
9 parameters

include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, stop, tool_choice, tools, verbosity

Provider / route
google-vertex/europe
1M contextunknown128K max output
Pricing
Input
$5.5 / 1M
Output
$27.5 / 1M
Cache
$0.55 / 1M
Blended— Worst comparable value
$12.8333 / 1M
Benchmarks
Speed
N/a
TTFT
N/a
TPS
N/a
Uptime
N/a
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls
9 parameters

include_reasoning, max_tokens, reasoning, reasoning_effort, response_format, stop, tool_choice, tools, verbosity

Green value Best comparable resultRed value Worst comparable result

Endpoint data fetched .

Claude Opus 5 deployment guide

How to choose a Claude Opus 5 inference provider

Claude Opus 5 is indexed here as a text model by anthropic, with 6 provider endpoints from Amazon Bedrock, Anthropic, Azure, Google Vertex. The comparison preserves each exact OpenRouter routing tag so pricing and performance observations can be connected to the route an application would actually request.

Match the endpoint to the workload

The model publishes a 1M-token context window. It accepts text, image, file input and returns text output. 12 distinct supported parameters appear across the listed routes. Confirm limits on the specific endpoint rather than assuming every host exposes the same configuration.

Compare the real request economics

Amazon Bedrock currently has the lowest estimated cost for the standard 1,000-input/500-output-token sample at $0.0175. Input-heavy and output-heavy applications can produce a different result, so review both per-million-token prices in the endpoint table.

Balance response start and generation speed

Azure currently has the shortest estimated 500-token response at 9.55 seconds. Azure is the single endpoint closest to the current ideal cost/speed combination. Recent observations can change, so validate finalists with your own prompts.