InferenceBench
Europe
Providers
Models
Research
Partners
Methodology
Search
/
DEV DATA
BENCHMARK
First benchmarks recorded: MiniMax M3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: Kimi K3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: DeepSeek V4 Pro → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-4.7 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-5.2 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: MiniMax M3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: Kimi K3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: DeepSeek V4 Pro → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-4.7 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-5.2 → Cortecs → ?
5h
Llama 3.3 70B
on
DeepInfra
#9 of 10 providers for Llama 3.3 70B · #3 of 3 models on DeepInfra · US
Compare with other providers
Score
69.7
Rel.
96.9%
Tools
95.5%
JSON
98.6%
TTFT
379 ms
P95
1212 ms
Tok/s
119
Price
$0.04 / $0.20
Ctx
128k
3,671 runs · last 27m ago ·
DEV DATA
Llama 3.3 70B on other providers
Provider
Score
TTFT
Tok/s
$ Out
Fireworks
91.9
231 ms
251
$0.48
Nebius
91.8
213 ms
192
$0.22
Groq
89.2
132 ms
1013
$0.32
Cerebras
88.5
110 ms
2835
$0.66
OpenRouter
76.8
284 ms
216
$0.48
Scaleway
75.1
263 ms
130
$0.40
Together
74.4
297 ms
176
$0.40
Nextbit
72.1
252 ms
113
$0.24
DeepInfra
THIS PAGE
69.7
379 ms
119
$0.20
OVHcloud
64.0
275 ms
100
$0.37
Other models on DeepInfra
Model
Score
TTFT
$ Out
Qwen3-235B
93.8
512 ms
$0.36
DeepSeek-V3
93.5
604 ms
$0.45
Llama 3.3 70B
THIS PAGE
93.3
379 ms
$0.20