Skip to content

Same-model provider benchmark

DeepInfra vs Mancer: LLM provider comparison

Compare DeepInfra and Mancer on 3 exact shared text models. ProviderBench keeps speed, price, and catalog coverage separate so naturally faster model catalogs cannot distort the result.

DeepInfra

85 indexed models

Headquarters
United States
Server regions
N/a
Model types
65 text, 15 embeddings, 5 speech

Mancer

6 indexed models

Headquarters
N/a
Server regions
N/a
Model types
6 text

At a glance

Metric winners

There is no overall score. Each winner answers one specific question using only directly comparable data.

Fastest on shared models

1.48× typical advantage

3 exact models with complete recent speed data

Lowest token cost

$0.2004 / 1M

3 exact models using a 1K-input/500-output mix

Most models available

85 models

Complete catalog coverage across all indexed modalities

Shared-model benchmark summary

MetricDeepInfra(0)Mancer(0)
SpeedDeepInfra11.37 sMancer21.66 s
TTFTDeepInfra0.36 sMancer0.83 s
TPSDeepInfra45.0 tok/sMancer24.0 tok/s
UptimeDeepInfra98.82%Mancer96.27%
BlendedDeepInfra$0.2004 / 1MMancer$0.4067 / 1M
Green value Better comparable resultRed value Worse comparable result

Visual comparison

Price and performance charts

DeepInfraMancer
500-token response by shared model

Estimated seconds using recent median response-start and output-speed data. Lower is better.

DeepInfra has the lower typical same-model response ratio across 3 measured models.

Blended token price by shared model

USD per 1 million tokens using a 1,000-input/500-output mix. Lower is better.

DeepInfra has the lower typical price ratio across 3 priced shared models.

Shared text models

3 exact models · newest first

ModelDeepInfra(0)Mancer(0)
DeepSeek V4 FlashWinner · DeepInfra
DeepInfra
Speed— Best comparable value
18.24 s
TTFT— Worst comparable value
1.00 s
TPS— Best comparable value
29.0 tok/s
Uptime— Best comparable value
98.82%
Context
1M
Route
deepinfra/fp4
Blended— Best comparable value
$0.12 / 1M
Mancer
Speed— Worst comparable value
21.66 s
TTFT— Best comparable value
0.83 s
TPS— Worst comparable value
24.0 tok/s
Uptime— Worst comparable value
96.27%
Context
1M
Route
mancer/fp4
Blended— Worst comparable value
$0.5 / 1M
gpt-oss-120bWinner · DeepInfra
DeepInfra
Speed— Best comparable value
10.57 s
TTFT— Best comparable value
0.36 s
TPS— Best comparable value
49.0 tok/s
Uptime— Best comparable value
87.51%
Context
131.1K
Route
deepinfra/bf16
Blended— Best comparable value
$0.0813 / 1M
Mancer
Speed— Worst comparable value
26.86 s
TTFT— Worst comparable value
1.86 s
TPS— Worst comparable value
20.0 tok/s
Uptime— Worst comparable value
71.47%
Context
131.1K
Route
mancer/fp8
Blended— Worst comparable value
$0.2033 / 1M
MythoMax 13BWinner · DeepInfra
DeepInfra
Speed— Best comparable value
11.37 s
TTFT— Best comparable value
0.26 s
TPS— Best comparable value
45.0 tok/s
Uptime— Best comparable value
100.00%
Context
4.1K
Route
deepinfra/fp16
Blended— Best comparable value
$0.4 / 1M
Mancer
Speed— Worst comparable value
16.79 s
TTFT— Worst comparable value
0.66 s
TPS— Worst comparable value
31.0 tok/s
Uptime— Worst comparable value
99.34%
Context
8.2K
Route
mancer/fp8
Blended— Worst comparable value
$0.5167 / 1M

A per-model winner combines blended price and estimated 500-token response time with equal proportional weight. Ties and rows missing either measurement receive no badge. Comparison data calculated . Values use one deterministic route per provider and model; missing measurements remain visible as N/a.

DeepInfra vs Mancer analysis

How DeepInfra and Mancer compare for AI inference

DeepInfra and Mancer share 3 indexed text models, including DeepSeek V4 Flash, gpt-oss-120b, MythoMax 13B. DeepInfra has the stronger typical response-time result on the directly measured set.

Same-model speed evidence

3 shared models currently have complete response-start and output-speed measurements on both providers. The speed comparison uses per-model ratios before taking the median, so naturally faster model catalogs do not improve the result.

Token pricing on one workload

3 shared models have complete input and output prices on both providers. Prices use the same 1,000-input/500-output-token mix and are normalized to one million tokens for readability.

Catalog and deployment differences

DeepInfra has 85 indexed models and lists no published server regions; Mancer has 6 models and lists no regions. Verify data residency, privacy terms, limits, and production latency directly before choosing.