PRICE~moonshotai/kimi-latest cached_input_per_mtok decreased from 0.80 to 0.2929mPRICE~moonshotai/kimi-latest output_per_mtok decreased from 13.00 to 11.3629mPRICE~moonshotai/kimi-latest input_per_mtok decreased from 1.39 to 1.0029mPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok increased from 0.95 to 1.0429mPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0058 to 0.003829mPRICEinclusionai/ling-3.0-flash-fin cached_input_per_mtok decreased from 0.012 to 0.008429mPRICEinclusionai/ling-3.0-flash-fin output_per_mtok decreased from 0.18 to 0.1229mPRICEinclusionai/ling-3.0-flash-fin input_per_mtok decreased from 0.06 to 0.04229mPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.0077 to 0.005129mPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.0077 to 0.005129mPRICE~z-ai/glm-flash-latest output_per_mtok increased from 0.60 to 0.931hPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.20 to 0.951hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0086 to 0.00581hPRICEz-ai/glm-5.2 cached_input_per_mtok decreased from 0.32 to 0.261hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.012 to 0.00771hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.012 to 0.00771hPRICE~moonshotai/kimi-latest input_per_mtok increased from 1.34 to 1.392hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.0038 to 0.00112hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok increased from 0.0038 to 0.00862hPRICEgoogle/gemma-4-26b-a4b-it cached_input_per_mtok decreased from 0.05 to 0.0382hPRICEgoogle/gemma-4-26b-a4b-it output_per_mtok decreased from 0.30 to 0.232hPRICEgoogle/gemma-4-26b-a4b-it input_per_mtok decreased from 0.09 to 0.0682hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.017 to 0.0122hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.017 to 0.0122hPRICE~z-ai/glm-flash-latest output_per_mtok decreased from 0.90 to 0.602hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.0058 to 0.00382hPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.47 to 1.202hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0058 to 0.00382hPRICEz-ai/glm-5.2 cached_input_per_mtok increased from 0.26 to 0.322hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok increased from 0.0077 to 0.0172hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok increased from 0.0077 to 0.0172hPRICE~moonshotai/kimi-latest cached_input_per_mtok increased from 0.29 to 0.803hPRICE~moonshotai/kimi-latest output_per_mtok increased from 11.36 to 13.003hPRICE~moonshotai/kimi-latest input_per_mtok increased from 0.50 to 1.343hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.013 to 0.00583hPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.60 to 1.473hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.013 to 0.00583hPRICE~deepseek/deepseek-flash-latest cached_input_per_mtok increased from 0.006 to 0.023hPRICE~deepseek/deepseek-flash-latest output_per_mtok increased from 0.50 to 0.603hPRICEz-ai/glm-5.3-flash cached_input_per_mtok increased from 0.011 to 0.033h

Scaleway vs Together

5 shared models · FR vs US · independent benchmarks

Registry facts

FactScalewayTogether
Typesovereign cloudinference host
HeadquartersFRUS
Catalogued endpoints11
EU residencyYesNo

GLM-5.2: Scaleway vs Together

● MEASURED Scaleway detail → · ● MEASURED Together detail →

MetricScalewayTogether
TTFT p50282 ms271 ms
TTFT p951630 ms589 ms
Throughput91 tok/s283 tok/s
Reliability100%100%
Tool calling100%100%
Structured output100%100%
$ output / M tokens$6.30$4.40

GPT-OSS 120B: Scaleway vs Together

● MEASURED Scaleway detail → · ● MEASURED Together detail →

MetricScalewayTogether
TTFT p50179 ms182 ms
TTFT p952214 ms467 ms
Throughput96 tok/s138 tok/s
Reliability100%93.5%
Tool calling100%100%
Structured output100%85.7%
$ output / M tokens$0.69$0.60

DeepSeek V4 Flash: Scaleway vs Together

● MEASURED Scaleway detail → · ● MEASURED Together detail →

MetricScalewayTogether
TTFT p50389 ms714 ms
TTFT p953314 ms1866 ms
Throughput133 tok/s61 tok/s
Reliability100%98.6%
Tool calling100%100%
Structured output100%100%
$ output / M tokens$0.92$0.28

Llama 3.3 70B: Scaleway vs Together

● MEASURED Scaleway detail → · ● MEASURED Together detail →

MetricScalewayTogether
TTFT p50144 ms618 ms
TTFT p95649 ms1761 ms
Throughput86 tok/s61 tok/s
Reliability100%100%
Tool calling100%98%
Structured output100%100%
$ output / M tokens$1.03$1.04

Every figure above is an independent measurement from the canonical vantage; unmeasured fields show — until measured. How we benchmark →