InferenceBench
Europe
Providers
Models
Research
Partners
Methodology
Search
/
DEV DATA
BENCHMARK
First benchmarks recorded: MiniMax M3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: Kimi K3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: DeepSeek V4 Pro → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-4.7 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-5.2 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: MiniMax M3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: Kimi K3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: DeepSeek V4 Pro → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-4.7 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-5.2 → Cortecs → ?
5h
Llama 3.3 70B
on
Nebius
#2 of 10 providers for Llama 3.3 70B · #3 of 3 models on Nebius · EU
Compare with other providers
Score
91.8
Rel.
98.8%
Tools
98.0%
JSON
99.3%
TTFT
213 ms
P95
778 ms
Tok/s
192
Price
$0.07 / $0.22
Ctx
128k
3,944 runs · last 15m ago ·
DEV DATA
Llama 3.3 70B on other providers
Provider
Score
TTFT
Tok/s
$ Out
Fireworks
91.9
231 ms
251
$0.48
Nebius
THIS PAGE
91.8
213 ms
192
$0.22
Groq
89.2
132 ms
1013
$0.32
Cerebras
88.5
110 ms
2835
$0.66
OpenRouter
76.8
284 ms
216
$0.48
Scaleway
75.1
263 ms
130
$0.40
Together
74.4
297 ms
176
$0.40
Nextbit
72.1
252 ms
113
$0.24
DeepInfra
69.7
379 ms
119
$0.20
OVHcloud
64.0
275 ms
100
$0.37
Other models on Nebius
Model
Score
TTFT
$ Out
Qwen3-235B
97.0
288 ms
$0.40
DeepSeek-V3
96.8
340 ms
$0.50
Llama 3.3 70B
THIS PAGE
95.9
213 ms
$0.22