InferenceBench
Europe
Providers
Models
Research
Partners
Methodology
Search
/
DEV DATA
BENCHMARK
First benchmarks recorded: MiniMax M3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: Kimi K3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: DeepSeek V4 Pro → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-4.7 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-5.2 → Cortecs → ?
6h
BENCHMARK
First benchmarks recorded: MiniMax M3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: Kimi K3 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: DeepSeek V4 Pro → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-4.7 → Cortecs → ?
5h
BENCHMARK
First benchmarks recorded: GLM-5.2 → Cortecs → ?
6h
Llama 3.3 70B
on
Cerebras
#4 of 10 providers for Llama 3.3 70B · #2 of 2 models on Cerebras · US
Compare with other providers
Score
88.5
Rel.
98.9%
Tools
98.2%
JSON
98.5%
TTFT
110 ms
P95
272 ms
Tok/s
2835
Price
$0.33 / $0.66
Ctx
128k
4,102 runs · last 12m ago ·
DEV DATA
Llama 3.3 70B on other providers
Provider
Score
TTFT
Tok/s
$ Out
Fireworks
91.9
231 ms
251
$0.48
Nebius
91.8
213 ms
192
$0.22
Groq
89.2
132 ms
1013
$0.32
Cerebras
THIS PAGE
88.5
110 ms
2835
$0.66
OpenRouter
76.8
284 ms
216
$0.48
Scaleway
75.1
263 ms
130
$0.40
Together
74.4
297 ms
176
$0.40
Nextbit
72.1
252 ms
113
$0.24
DeepInfra
69.7
379 ms
119
$0.20
OVHcloud
64.0
275 ms
100
$0.37
Other models on Cerebras
Model
Score
TTFT
$ Out
Qwen3-235B
97.0
148 ms
$1.20
Llama 3.3 70B
THIS PAGE
96.2
110 ms
$0.66