Skip to content

Highest Output-Speed Laguna S 2.1 Inference Providers

Ranks providers by their recent typical output speed in tokens per second. Providers without enough recent data remain visible after ranked rows.

Median input price:
$0.1 / 1M
Median output price:
$0.2 / 1M
Median cache price:
$0.01 / 1M

Highest output speed endpoint ranking

Highest Output-Speed Laguna S 2.1 Inference Providers
RankProvider / routePricingBenchmarksAPI support
#1Provider / route
1M contextbf16131.1K max output
Pricing
Input
$0.1 / 1M
Output
$0.2 / 1M
Cache
$0.01 / 1M
Blended
$0.1333 / 1M
Benchmarks
Speed
8.39 s
TTFT
0.58 s
TPS
64.0 tok/s
Uptime
100.00%
API support
  • Tool calling
  • Tool choice
  • Structured output
  • Parallel calls

Endpoint data fetched .

Laguna S 2.1 endpoint guide

How to interpret the Highest Output-Speed Laguna S 2.1 Inference Providers

1 of 1 Laguna S 2.1 endpoint currently have the published data required for this ranking. Poolside leads at 64.0 tok/s via poolside/bf16. The table keeps unranked routes visible so missing measurements do not look like missing provider availability.

How generation throughput is ranked

The output-speed ranking compares recent median generated tokens per second. It favors routes that complete long generations quickly, while first-token latency remains a separate measure of how responsive the endpoint feels at the beginning.

Compare Poolside with the next option

Only 1 route currently qualifies, so the rank alone provides limited choice. Review every visible endpoint for price, speed, uptime, context length, and route-specific configuration.

Use the route, not only the provider name

Laguna S 2.1 has 1 published provider option, and performance data is matched to each exact OpenRouter routing tag. Different quantization, context, regional deployment, or provider configuration can change price and behavior even when the underlying model name is identical.