Skip to content

Same-model provider benchmark

Baidu Qianfan vs Cloudflare: LLM provider comparison

Compare Baidu Qianfan and Cloudflare on 3 exact shared text models. ProviderBench keeps speed, price, and catalog coverage separate so naturally faster model catalogs cannot distort the result.

Baidu Qianfan

7 indexed models

Headquarters
China
Server regions
N/a
Model types
7 text

Cloudflare

13 indexed models

Headquarters
United States
Server regions
N/a
Model types
13 text

At a glance

Metric winners

There is no overall score. Each winner answers one specific question using only directly comparable data.

Fastest on shared models

1.36× typical advantage

3 exact models with complete recent speed data

Lowest token cost

$1.2957 / 1M

3 exact models using a 1K-input/500-output mix

Most models available

13 models

Complete catalog coverage across all indexed modalities

Shared-model benchmark summary

MetricBaidu Qianfan(0)Cloudflare(0)
SpeedBaidu Qianfan9.65 sCloudflare12.82 s
TTFTBaidu Qianfan0.88 sCloudflare1.01 s
TPSBaidu Qianfan57.0 tok/sCloudflare42.0 tok/s
UptimeBaidu Qianfan99.80%Cloudflare100.00%
BlendedBaidu Qianfan$1.2957 / 1MCloudflare$1.5178 / 1M
Green value Better comparable resultRed value Worse comparable result

Visual comparison

Price and performance charts

Baidu QianfanCloudflare
500-token response by shared model

Estimated seconds using recent median response-start and output-speed data. Lower is better.

Cloudflare has the lower typical same-model response ratio across 3 measured models.

Blended token price by shared model

USD per 1 million tokens using a 1,000-input/500-output mix. Lower is better.

Baidu Qianfan has the lower typical price ratio across 3 priced shared models.

Shared text models

3 exact models · newest first

ModelBaidu Qianfan(0)Cloudflare(0)
GLM 5.2Winner · Cloudflare
Baidu Qianfan
Speed— Worst comparable value
9.65 s
TTFT— Best comparable value
0.88 s
TPS— Worst comparable value
57.0 tok/s
Uptime— Worst comparable value
99.64%
Context
1M
Route
baidu/fp8
Blended— Best comparable value
$2.16 / 1M
Cloudflare
Speed— Best comparable value
7.12 s
TTFT— Worst comparable value
1.02 s
TPS— Best comparable value
82.0 tok/s
Uptime— Best comparable value
100.00%
Context
262.1K
Route
cloudflare
Blended— Worst comparable value
$2.4 / 1M
DeepSeek V4 FlashWinner · Baidu Qianfan
Baidu Qianfan
Speed— Best comparable value
7.58 s
TTFT— Best comparable value
0.63 s
TPS— Best comparable value
72.0 tok/s
Uptime— Worst comparable value
99.95%
Context
1M
Route
baidu/fp8
Blended— Best comparable value
$0.1311 / 1M
Cloudflare
Speed— Worst comparable value
17.14 s
TTFT— Worst comparable value
1.01 s
TPS— Worst comparable value
31.0 tok/s
Uptime— Best comparable value
100.00%
Context
384K
Route
cloudflare
Blended— Worst comparable value
$0.1867 / 1M
Kimi K2.6Winner · Cloudflare
Baidu Qianfan
Speed— Worst comparable value
20.07 s
TTFT— Worst comparable value
1.55 s
TPS— Worst comparable value
27.0 tok/s
Uptime
99.98%
Context
262.1K
Route
baidu/fp4
Blended— Best comparable value
$1.596 / 1M
Cloudflare
Speed— Best comparable value
12.82 s
TTFT— Best comparable value
0.92 s
TPS— Best comparable value
42.0 tok/s
Uptime
N/a
Context
262.1K
Route
cloudflare
Blended— Worst comparable value
$1.9667 / 1M

A per-model winner combines blended price and estimated 500-token response time with equal proportional weight. Ties and rows missing either measurement receive no badge. Comparison data calculated . Values use one deterministic route per provider and model; missing measurements remain visible as N/a.

Baidu Qianfan vs Cloudflare analysis

How Baidu Qianfan and Cloudflare compare for AI inference

Baidu Qianfan and Cloudflare share 3 indexed text models, including GLM 5.2, DeepSeek V4 Flash, Kimi K2.6. Cloudflare has the stronger typical response-time result on the directly measured set.

Same-model speed evidence

3 shared models currently have complete response-start and output-speed measurements on both providers. The speed comparison uses per-model ratios before taking the median, so naturally faster model catalogs do not improve the result.

Token pricing on one workload

3 shared models have complete input and output prices on both providers. Prices use the same 1,000-input/500-output-token mix and are normalized to one million tokens for readability.

Catalog and deployment differences

Baidu Qianfan has 7 indexed models and lists no published server regions; Cloudflare has 13 models and lists no regions. Verify data residency, privacy terms, limits, and production latency directly before choosing.