Skip to content

Same-model provider benchmark

Baseten vs CoreWeave: LLM provider comparison

Compare Baseten and CoreWeave on 5 exact shared text models. ProviderBench keeps speed, price, and catalog coverage separate so naturally faster model catalogs cannot distort the result.

Baseten

7 indexed models

Headquarters
United States
Server regions
N/a
Model types
7 text

CoreWeave

20 indexed models

Headquarters
United States
Server regions
United States
Model types
20 text

At a glance

Metric winners

There is no overall score. Each winner answers one specific question using only directly comparable data.

Fastest on shared models

1.98× typical advantage

5 exact models with complete recent speed data

Lowest token cost

$1.8307 / 1M

5 exact models using a 1K-input/500-output mix

Most models available

20 models

Complete catalog coverage across all indexed modalities

Shared-model benchmark summary

MetricBaseten(0)CoreWeave(0)
SpeedBaseten8.69 sCoreWeave5.92 s
TTFTBaseten1.22 sCoreWeave0.50 s
TPSBaseten66.0 tok/sCoreWeave90.0 tok/s
UptimeBaseten99.10%CoreWeave99.82%
BlendedBaseten$1.844 / 1MCoreWeave$1.8307 / 1M
Green value Better comparable resultRed value Worse comparable result

Visual comparison

Price and performance charts

BasetenCoreWeave
500-token response by shared model

Estimated seconds using recent median response-start and output-speed data. Lower is better.

CoreWeave has the lower typical same-model response ratio across 5 measured models.

Blended token price by shared model

USD per 1 million tokens using a 1,000-input/500-output mix. Lower is better.

CoreWeave has the lower typical price ratio across 5 priced shared models.

Shared text models

5 exact models · newest first

ModelBaseten(0)CoreWeave(0)
GLM 5.2Winner · Baseten
Baseten
Speed— Best comparable value
13.22 s
TTFT— Worst comparable value
2.11 s
TPS— Best comparable value
45.0 tok/s
Uptime— Best comparable value
100.00%
Context
524.3K
Route
baseten/fp8
Blended— Worst comparable value
$2.4 / 1M
CoreWeave
Speed— Worst comparable value
20.46 s
TTFT— Best comparable value
1.94 s
TPS— Worst comparable value
27.0 tok/s
Uptime— Worst comparable value
99.96%
Context
262.1K
Route
coreweave/fp4
Blended— Best comparable value
$2.3933 / 1M
DeepSeek V4 ProWinner · Baseten
Baseten
Speed— Best comparable value
8.38 s
TTFT— Best comparable value
0.81 s
TPS— Best comparable value
66.0 tok/s
Uptime— Worst comparable value
98.46%
Context
262.1K
Route
baseten/fp4
Blended
$2.32 / 1M
CoreWeave
Speed— Worst comparable value
64.00 s
TTFT— Worst comparable value
1.50 s
TPS— Worst comparable value
8.0 tok/s
Uptime— Best comparable value
98.59%
Context
1M
Route
coreweave/fp8
Blended
$2.32 / 1M
Kimi K2.6Winner · CoreWeave
Baseten
Speed— Worst comparable value
6.71 s
TTFT— Worst comparable value
2.77 s
TPS— Worst comparable value
127.0 tok/s
Uptime
N/a
Context
262K
Route
baseten/fp4
Blended
$1.9667 / 1M
CoreWeave
Speed— Best comparable value
3.30 s
TTFT— Best comparable value
0.34 s
TPS— Best comparable value
169.0 tok/s
Uptime
99.89%
Context
262.1K
Route
coreweave/fp4
Blended
$1.9667 / 1M
GLM 5.1Winner · CoreWeave
Baseten
Speed— Worst comparable value
8.69 s
TTFT— Worst comparable value
1.22 s
TPS— Worst comparable value
67.0 tok/s
Uptime— Worst comparable value
89.84%
Context
202.8K
Route
baseten/fp4
Blended— Best comparable value
$2.3 / 1M
CoreWeave
Speed— Best comparable value
4.38 s
TTFT— Best comparable value
0.50 s
TPS— Best comparable value
129.0 tok/s
Uptime— Best comparable value
100.00%
Context
202.8K
Route
coreweave/fp8
Blended— Worst comparable value
$2.4 / 1M
gpt-oss-120bWinner · CoreWeave
Baseten
Speed— Worst comparable value
15.61 s
TTFT— Worst comparable value
0.46 s
TPS— Worst comparable value
33.0 tok/s
Uptime— Best comparable value
99.73%
Context
128.1K
Route
baseten/fp4
Blended— Worst comparable value
$0.2333 / 1M
CoreWeave
Speed— Best comparable value
5.92 s
TTFT— Best comparable value
0.36 s
TPS— Best comparable value
90.0 tok/s
Uptime— Worst comparable value
99.68%
Context
131.1K
Route
coreweave/fp4
Blended— Best comparable value
$0.0733 / 1M

A per-model winner combines blended price and estimated 500-token response time with equal proportional weight. Ties and rows missing either measurement receive no badge. Comparison data calculated . Values use one deterministic route per provider and model; missing measurements remain visible as N/a.

Baseten vs CoreWeave analysis

How Baseten and CoreWeave compare for AI inference

Baseten and CoreWeave share 5 indexed text models, including GLM 5.2, DeepSeek V4 Pro, Kimi K2.6, GLM 5.1, gpt-oss-120b. CoreWeave has the stronger typical response-time result on the directly measured set.

Same-model speed evidence

5 shared models currently have complete response-start and output-speed measurements on both providers. The speed comparison uses per-model ratios before taking the median, so naturally faster model catalogs do not improve the result.

Token pricing on one workload

5 shared models have complete input and output prices on both providers. Prices use the same 1,000-input/500-output-token mix and are normalized to one million tokens for readability.

Catalog and deployment differences

Baseten has 7 indexed models and lists no published server regions; CoreWeave has 20 models and lists 1 region. Verify data residency, privacy terms, limits, and production latency directly before choosing.