Nebius
14 indexed models
- Headquarters
Netherlands
- Server regions
- N/a
- Model types
- 13 text, 1 embeddings
Start typing to search.
Same-model provider benchmark
Compare Nebius and Together on 3 exact shared text models. ProviderBench keeps speed, price, and catalog coverage separate so naturally faster model catalogs cannot distort the result.
14 indexed models
23 indexed models
At a glance
There is no overall score. Each winner answers one specific question using only directly comparable data.
| Metric | ||
|---|---|---|
| Speed | Nebius11.76 s | Together16.85 s |
| TTFT | Nebius0.83 s | Together1.75 s |
| TPS | Nebius46.0 tok/s | Together37.0 tok/s |
| Uptime | Nebius94.36% | Together94.97% |
| Blended | Nebius$2.5067 / 1M | Together$2.78 / 1M |
Visual comparison
Estimated seconds using recent median response-start and output-speed data. Lower is better.
Nebius has the lower typical same-model response ratio across 3 measured models.
USD per 1 million tokens using a 1,000-input/500-output mix. Lower is better.
Nebius has the lower typical price ratio across 3 priced shared models.
| Model | ||
|---|---|---|
Kimi K3Winner · Nebius | Nebius
| Together
|
gpt-oss-120bWinner · Nebius | Nebius
| Together
|
Llama 3.3 70B InstructWinner · Nebius | Nebius
| Together
|
A per-model winner combines blended price and estimated 500-token response time with equal proportional weight. Ties and rows missing either measurement receive no badge. Comparison data calculated . Values use one deterministic route per provider and model; missing measurements remain visible as N/a.
Nebius vs Together analysis
Nebius and Together share 3 indexed text models, including Kimi K3, gpt-oss-120b, Llama 3.3 70B Instruct. Nebius has the stronger typical response-time result on the directly measured set.
3 shared models currently have complete response-start and output-speed measurements on both providers. The speed comparison uses per-model ratios before taking the median, so naturally faster model catalogs do not improve the result.
3 shared models have complete input and output prices on both providers. Prices use the same 1,000-input/500-output-token mix and are normalized to one million tokens for readability.
Nebius has 14 indexed models and lists no published server regions; Together has 23 models and lists no regions. Verify data residency, privacy terms, limits, and production latency directly before choosing.
Related same-model benchmarks
Explore qualified alternatives with the most shared measured models. Recommendations include comparisons for both Nebius and Together.