PERFORMANCEGLM-5.2 → Scaleway: TTFT ↑ 20%28mPERFORMANCEDeepSeek V4 Flash → Scaleway: TTFT ↑ 35%28mPERFORMANCEMistral Medium 3.5 → Scaleway: TTFT ↓ 21%28mPERFORMANCEDeepSeek V4 Flash → Scaleway: throughput ↓ 20%28mPERFORMANCEGPT-OSS 120B → Scaleway: throughput ↑ 31%28mPERFORMANCEGPT-OSS 120B → Scaleway: TTFT ↓ 24%28mPERFORMANCEGPT-OSS 120B → Together AI: throughput ↑ 20%28mPERFORMANCEKimi K3 via OpenRouter: TTFT ↓ 49%28mPERFORMANCEGLM-5.2 via OpenRouter: reliability recovered 96.8% → 100.0%28mPERFORMANCEKimi K3 via Cortecs: throughput ↑ 184%28mPERFORMANCEKimi K3 via Cortecs: TTFT ↓ 42%28mPERFORMANCEGLM-5.2 via OpenRouter: throughput ↓ 49%28mPERFORMANCEGLM-5.2 via OpenRouter: TTFT ↑ 121%28mPERFORMANCEMiniMax M3 via OpenRouter: reliability recovered 96.8% → 100.0%28mPERFORMANCEDeepSeek V4 Pro via Cortecs: throughput ↑ 27%28mPERFORMANCEGPT-OSS 120B via OpenRouter: throughput ↓ 22%28mPERFORMANCEGLM-4.7 via OpenRouter: TTFT ↓ 78%28mPERFORMANCEGPT-OSS 20B → Groq: TTFT ↑ 24%28mPERFORMANCEKimi K3 → Together AI: TTFT ↓ 34%28mPERFORMANCEDeepSeek V4 Flash via OpenRouter: throughput ↓ 17%28mPERFORMANCELlama 3.3 70B via Cortecs: TTFT ↓ 51%28mPERFORMANCEQwen3 235B via OpenRouter: throughput ↑ 37%28mPERFORMANCEQwen3 235B via OpenRouter: TTFT ↓ 29%28mPERFORMANCEDeepSeek V4 Flash → Together AI: throughput ↑ 19%28mPERFORMANCEGLM-5.3 via OpenRouter: throughput ↓ 17%28mPERFORMANCEGPT-OSS 20B via Cortecs: throughput ↑ 28%28mPERFORMANCELlama 3.3 70B → Together AI: TTFT ↓ 33%28mPRICE~moonshotai/kimi-latest cached_input_per_mtok decreased from 0.80 to 0.2929mPRICE~moonshotai/kimi-latest output_per_mtok decreased from 13.00 to 11.3629mPRICE~moonshotai/kimi-latest input_per_mtok decreased from 0.99 to 0.7529mPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.0077 to 0.001129mPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.28 to 1.0429mPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0077 to 0.003829mPRICEz-ai/glm-5.1 cached_input_per_mtok increased from 0.18 to 0.2629mPRICEz-ai/glm-5.1 output_per_mtok increased from 3.03 to 4.4029mPRICEz-ai/glm-5.1 input_per_mtok increased from 0.96 to 1.4029mPRICEtencent/hy4-preview cached_input_per_mtok increased from 0.038 to 0.04229mPRICEtencent/hy4-preview input_per_mtok increased from 0.75 to 0.8329mPRICEtencent/hy3 cached_input_per_mtok increased from 0.021 to 0.03329mPRICEtencent/hy3 output_per_mtok increased from 0.33 to 0.5329m

Scaleway — AI inference benchmarks

FR · serverless · 6 benchmarked models · 1 region · EU residency
#4 GLOBALAdd to compare
Overall92.4Rel.100.0%Tools100.0%JSON100.0%TTFT172 msP951234 msTok/s93Price$0.95 / $1.81Δ30d↑ 4.8● last run 13h ago · 24 runs in 24h

Head-to-head: vs Cerebras · vs Groq · vs OpenRouter · vs Cortecs · vs Together · vs Mistral

Provider vs market

Percentile across 8 ranked providers
Reliability 100.0%P100
Tools 100.0%P100
Structured 100.0%P100
TTFT 172 msP100
Throughput 93 tok/sP25
Cost $1.81P38

Price vs performance

Scaleway highlighted vs peers
889296100$4.37$8.74$13.10$17.47Output price $/M →Score ↑Groq — $0.45/M out · score 99.2Cerebras — $0.75/M out · score 99.0Mistral — $4.50/M out · score 93.2Scaleway — $1.81/M out · score 92.4ScalewayCortecs — $1.23/M out · score 92.1OpenAI — $15.60/M out · score 90.4Together — $1.12/M out · score 85.5OpenRouter — $1.20/M out · score 85.2

30-day trend

Measured daily series
Score92.4
TTFT162 ms
Tools100.0%
Reliability100.0%
Price$2.58

Key positioning

Derived from benchmark percentiles
Strengths
  • Reliability (P100)
  • Tools (P100)
  • Structured (P100)
  • TTFT (P100)
Weaknesses
  • Throughput (P25)
  • Cost (P38)
Best suited for

Scaleway in the registry

  • Parent companyIliad Group
  • HeadquartersFR
  • API protocolsopenai chat completions
  • Catalogued endpoints1
  • Websitewww.scaleway.com
  • Last verified2026-10-03

Observed models & declared pricing

Read from Scaleway's own catalogue by continuous discovery — declared data, not benchmark results.

ModelObserved
bge-multilingual-gemma22026-10-03
deepseek-v4-flash-07312026-10-03
gemma-4-26b-a4b-it2026-10-03
glm-5.22026-10-03
gpt-oss-120b2026-10-03
llama-3.3-70b-instruct2026-10-03
mistral-medium-3.5-128b2026-10-03
mistral-small-3.2-24b-instruct-25062026-10-03
pixtral-12b-24092026-10-03
qwen3-235b-a22b-instruct-25072026-10-03
qwen3-coder-30b-a3b-instruct2026-10-03
qwen3-embedding-8b2026-10-03
qwen3.5-397b-a17b2026-10-03
qwen3.6-35b-a3b2026-10-03
qwen3.8-27b2026-10-03
whisper-large-v32026-10-03

All providers · How we benchmark