Skip to content

Cheapest LongCat 2.0 Inference Providers

Ranks providers by estimated cost using a 1,000-input/500-output token ratio, shown per 1 million total tokens.

Median input price:
$0.3 / 1M
Median output price:
$1.2 / 1M
Median cache price:
$0.006 / 1M

Cheapest endpoint ranking

Cheapest LongCat 2.0 Inference Providers
RankProvider / routePricingBenchmarksAPI support
#1Provider / route
1M contextfp8262.1K max output
Pricing
Input
$0.3 / 1M
Output
$1.2 / 1M
Cache
$0.006 / 1M
Blended
$0.6 / 1M
Benchmarks
Speed
20.64 s
TTFT
2.79 s
TPS
28.0 tok/s
Uptime
100.00%
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls

Endpoint data fetched .

LongCat 2.0 endpoint guide

How to interpret the Cheapest LongCat 2.0 Inference Providers

1 of 1 LongCat 2.0 endpoint currently have the published data required for this ranking. AtlasCloud leads at $0.6 / 1M via atlas-cloud/fp8. The table keeps unranked routes visible so missing measurements do not look like missing provider availability.

How estimated token cost is ranked

The cheapest ranking uses a 1,000-input/500-output-token ratio and reports the blended result per 1 million total tokens. Because providers can price input and output differently, a workload with a different token ratio may produce a different order.

Compare AtlasCloud with the next option

Only 1 route currently qualifies, so the rank alone provides limited choice. Review every visible endpoint for price, speed, uptime, context length, and route-specific configuration.

Use the route, not only the provider name

LongCat 2.0 has 1 published provider option, and performance data is matched to each exact OpenRouter routing tag. Different quantization, context, regional deployment, or provider configuration can change price and behavior even when the underlying model name is identical.