Skip to content

Meituan: LongCat 2.0 provider comparison

meituan/longcat-2.0

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...

Endpoints:
1
Context:
1M
Input:
text
Output:
text
Compare this model in a set →

At a glance

Comparison winners

Based on 1,000 input and 500 output tokens using the latest published median speed data.

Best cost/speed trade-off

AtlasCloudatlas-cloud/fp8

$0.6 / 1M tokens

17.95 s estimated response

Cheapest

AtlasCloudatlas-cloud/fp8

$0.6 / 1M tokens

$0.0009 for the sample request

Fastest

AtlasCloudatlas-cloud/fp8

17.95 s estimated response

$0.6 / 1M tokens

Visual comparison

Price and performance charts

Compare published endpoint pricing and estimated response time visually.

Effective price by endpoint

USD per 1 million tokens for input and output. Lower is better.

AtlasCloud has the lowest estimated cost for 1,000 input and 500 output tokens at $0.0009.

Estimated cost vs. response time

Based on 1,000 input and 500 output tokens. Cost is shown in USD per 1 million tokens. Lower and further left is better; the single best cost/speed trade-off appears at full opacity.

1 provider has complete pricing and speed data. AtlasCloud is closest to the ideal combination of lowest cost and fastest response.

Server-rendered comparison

Provider endpoints

Prompt and completion prices are effective USD per token as published by OpenRouter. Missing speed data never removes an endpoint.

Endpoint comparison for Meituan: LongCat 2.0
Provider / routePricing ↗BenchmarksAPI support
Provider / route
atlas-cloud/fp8
1M contextfp8262.1K max output
CheapestFastestBest cost/speed trade-off
Pricing
Input
$0.3 / 1M
Output
$1.2 / 1M
Cache
$0.006 / 1M
Blended
$0.6 / 1M
Discount
60.0%
Benchmarks
Speed
17.95 s
TTFT
2.80 s
TPS
33.0 tok/s
Uptime
100.00%
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls
15 parameters

frequency_penalty, include_reasoning, logit_bias, max_tokens, min_p, presence_penalty, reasoning, repetition_penalty, seed, stop, temperature, tool_choice, tools, top_k, top_p

Endpoint data fetched .

LongCat 2.0 deployment guide

How to choose a LongCat 2.0 inference provider

LongCat 2.0 is indexed here as a text model by Meituan, with 1 provider endpoint from AtlasCloud. The comparison preserves each exact OpenRouter routing tag so pricing and performance observations can be connected to the route an application would actually request.

Match the endpoint to the workload

The model publishes a 1M-token context window. It accepts text input and returns text output. 15 distinct supported parameters appear across the listed routes. Confirm limits on the specific endpoint rather than assuming every host exposes the same configuration.

Compare the real request economics

AtlasCloud currently has the lowest estimated cost for the standard 1,000-input/500-output-token sample at $0.0009. Input-heavy and output-heavy applications can produce a different result, so review both per-million-token prices in the endpoint table.

Balance response start and generation speed

AtlasCloud currently has the shortest estimated 500-token response at 17.95 seconds. AtlasCloud is the single endpoint closest to the current ideal cost/speed combination. Recent observations can change, so validate finalists with your own prompts.