PRICE~moonshotai/kimi-latest cached_input_per_mtok decreased from 0.80 to 0.298mPRICE~moonshotai/kimi-latest output_per_mtok decreased from 13.00 to 11.368mPRICE~moonshotai/kimi-latest input_per_mtok decreased from 1.39 to 1.008mPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok increased from 0.95 to 1.048mPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0058 to 0.00388mPRICEinclusionai/ling-3.0-flash-fin cached_input_per_mtok decreased from 0.012 to 0.00848mPRICEinclusionai/ling-3.0-flash-fin output_per_mtok decreased from 0.18 to 0.128mPRICEinclusionai/ling-3.0-flash-fin input_per_mtok decreased from 0.06 to 0.0428mPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.0077 to 0.00518mPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.0077 to 0.00518mPRICE~z-ai/glm-flash-latest output_per_mtok increased from 0.60 to 0.9339mPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.20 to 0.9539mPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0086 to 0.005839mPRICEz-ai/glm-5.2 cached_input_per_mtok decreased from 0.32 to 0.2639mPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.012 to 0.007739mPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.012 to 0.007739mPRICE~moonshotai/kimi-latest input_per_mtok increased from 1.34 to 1.391hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.0038 to 0.00111hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok increased from 0.0038 to 0.00861hPRICEgoogle/gemma-4-26b-a4b-it cached_input_per_mtok decreased from 0.05 to 0.0381hPRICEgoogle/gemma-4-26b-a4b-it output_per_mtok decreased from 0.30 to 0.231hPRICEgoogle/gemma-4-26b-a4b-it input_per_mtok decreased from 0.09 to 0.0681hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok decreased from 0.017 to 0.0121hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok decreased from 0.017 to 0.0121hPRICE~z-ai/glm-flash-latest output_per_mtok decreased from 0.90 to 0.602hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.0058 to 0.00382hPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.47 to 1.202hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.0058 to 0.00382hPRICEz-ai/glm-5.2 cached_input_per_mtok increased from 0.26 to 0.322hPRICEdeepseek/deepseek-v4-flash-0731 cached_input_per_mtok increased from 0.0077 to 0.0172hPRICEdeepseek/deepseek-v4-flash-0731 input_per_mtok increased from 0.0077 to 0.0172hPRICE~moonshotai/kimi-latest cached_input_per_mtok increased from 0.29 to 0.802hPRICE~moonshotai/kimi-latest output_per_mtok increased from 11.36 to 13.002hPRICE~moonshotai/kimi-latest input_per_mtok increased from 0.50 to 1.342hPRICE~deepseek/deepseek-v4-flash-latest cached_input_per_mtok decreased from 0.013 to 0.00582hPRICE~deepseek/deepseek-v4-flash-latest output_per_mtok decreased from 1.60 to 1.472hPRICE~deepseek/deepseek-v4-flash-latest input_per_mtok decreased from 0.013 to 0.00582hPRICE~deepseek/deepseek-flash-latest cached_input_per_mtok increased from 0.006 to 0.022hPRICE~deepseek/deepseek-flash-latest output_per_mtok increased from 0.50 to 0.602hPRICEz-ai/glm-5.3-flash cached_input_per_mtok increased from 0.011 to 0.032h

Groq vs Scaleway

1 shared model · US vs FR · independent benchmarks

Registry facts

FactGroqScaleway
Typeinference hostsovereign cloud
HeadquartersUSFR
Catalogued endpoints11
EU residencyNoYes

GPT-OSS 120B: Groq vs Scaleway

● MEASURED Groq detail → · ● MEASURED Scaleway detail →

MetricGroqScaleway
TTFT p50264 ms179 ms
TTFT p95402 ms2214 ms
Throughput486 tok/s96 tok/s
Reliability100%100%
Tool calling100%100%
Structured output100%100%
$ output / M tokens$0.60$0.69

Every figure above is an independent measurement from the canonical vantage; unmeasured fields show — until measured. How we benchmark →