PRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.04 to 0.6017mPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok increased from 0.0038 to 0.01317mPRICEnvidia/nemotron-3-ultra-550b-a55b cached_input_per_mtok decreased from 0.12 to 0.1017mPRICEnvidia/nemotron-3-ultra-550b-a55b output_per_mtok decreased from 2.40 to 2.2017mPRICEnvidia/nemotron-3-ultra-550b-a55b input_per_mtok decreased from 0.60 to 0.5017mPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok increased from 0.0051 to 0.01717mPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok increased from 0.0051 to 0.01717mPRICE~moonshotai/kimi-latest cached_input_per_mtok decreased from 0.80 to 0.2948mPRICE~moonshotai/kimi-latest output_per_mtok decreased from 13.00 to 11.3648mPRICE~moonshotai/kimi-latest input_per_mtok decreased from 1.39 to 1.0048mPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok increased from 0.95 to 1.0448mPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0058 to 0.003848mPRICEinclusionai/ling-3.0-flash-fin cached_input_per_mtok decreased from 0.012 to 0.008448mPRICEinclusionai/ling-3.0-flash-fin output_per_mtok decreased from 0.18 to 0.1248mPRICEinclusionai/ling-3.0-flash-fin input_per_mtok decreased from 0.06 to 0.04248mPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.0077 to 0.005148mPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.0077 to 0.005148mPRICE~z-ai/glm-flash-latest output_per_mtok increased from 0.60 to 0.931hPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.20 to 0.951hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0086 to 0.00581hPRICEz-ai/glm-5.2 cached_input_per_mtok decreased from 0.32 to 0.261hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.012 to 0.00771hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.012 to 0.00771hPRICE~moonshotai/kimi-latest input_per_mtok increased from 1.34 to 1.392hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.0038 to 0.00112hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok increased from 0.0038 to 0.00862hPRICEgoogle/gemma-4-26b-a4b-it cached_input_per_mtok decreased from 0.05 to 0.0382hPRICEgoogle/gemma-4-26b-a4b-it output_per_mtok decreased from 0.30 to 0.232hPRICEgoogle/gemma-4-26b-a4b-it input_per_mtok decreased from 0.09 to 0.0682hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.017 to 0.0122hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.017 to 0.0122hPRICE~z-ai/glm-flash-latest output_per_mtok decreased from 0.90 to 0.602hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.0058 to 0.00382hPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.47 to 1.202hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0058 to 0.00382hPRICEz-ai/glm-5.2 cached_input_per_mtok increased from 0.26 to 0.322hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok increased from 0.0077 to 0.0172hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok increased from 0.0077 to 0.0172hPRICE~moonshotai/kimi-latest cached_input_per_mtok increased from 0.29 to 0.803hPRICE~moonshotai/kimi-latest output_per_mtok increased from 11.36 to 13.003h

Cortecs vs Scaleway

6 shared models · AT vs FR · independent benchmarks

Registry facts

FactCortecsScaleway
Typeai gatewaysovereign cloud
HeadquartersATFR
Catalogued endpoints11
EU residencyYesYes

GLM-5.2: Cortecs vs Scaleway

● MEASURED Cortecs detail → · ● MEASURED Scaleway detail →

MetricCortecsScaleway
TTFT p50425 ms282 ms
TTFT p952067 ms1630 ms
Throughput142 tok/s91 tok/s
Reliability100%100%
Tool calling100%100%
Structured output100%100%
$ output / M tokens$3.33$6.30

DeepSeek V4 Flash: Cortecs vs Scaleway

● MEASURED Cortecs detail → · ● MEASURED Scaleway detail →

MetricCortecsScaleway
TTFT p50411 ms389 ms
TTFT p951263 ms3314 ms
Throughput73 tok/s133 tok/s
Reliability100%100%
Tool calling100%100%
Structured output100%100%
$ output / M tokens$0.18$0.92

GPT-OSS 120B: Cortecs vs Scaleway

● MEASURED Cortecs detail → · ● MEASURED Scaleway detail →

MetricCortecsScaleway
TTFT p50240 ms179 ms
TTFT p95501 ms2214 ms
Throughput358 tok/s96 tok/s
Reliability100%100%
Tool calling100%100%
Structured output100%100%
$ output / M tokens$0.46$0.69

Qwen3-235B: Cortecs vs Scaleway

● MEASURED Cortecs detail → · ● MEASURED Scaleway detail →

MetricCortecsScaleway
TTFT p50937 ms162 ms
TTFT p955613 ms837 ms
Throughput52 tok/s81 tok/s
Reliability99.5%100%
Tool calling98%100%
Structured output100%100%
$ output / M tokens$0.47$2.58

Llama 3.3 70B: Cortecs vs Scaleway

● MEASURED Cortecs detail → · ● MEASURED Scaleway detail →

MetricCortecsScaleway
TTFT p50441 ms144 ms
TTFT p951051 ms649 ms
Throughput51 tok/s86 tok/s
Reliability100%100%
Tool calling98%100%
Structured output100%100%
$ output / M tokens$0.75$1.03

Mistral Medium 3.5: Cortecs vs Scaleway

● MEASURED Cortecs detail → · ● MEASURED Scaleway detail →

MetricCortecsScaleway
TTFT p50261 ms164 ms
TTFT p95441 ms597 ms
Throughput153 tok/s69 tok/s
Reliability99.5%100%
Tool calling100%100%
Structured output100%100%
$ output / M tokens$8.07$8.60

Every figure above is an independent measurement from the canonical vantage; unmeasured fields show — until measured. How we benchmark →