Head-to-head: vs Cerebras · vs OpenRouter · vs Cortecs · vs Together · vs Scaleway
Provider vs market
Percentile across 8 ranked providersReliability 100.0%P100
Tools 100.0%P100
Structured 100.0%P100
TTFT 261 msP75
Throughput 708 tok/sP88
Cost $0.45P100
Price vs performance
Groq highlighted vs peers30-day trend
Measured daily seriesKey positioning
Derived from benchmark percentilesStrengths- Reliability (P100)
- Tools (P100)
- Structured (P100)
- TTFT (P75)
- Throughput (P88)
- Cost (P100)
Groq in the registry
- HeadquartersUS
- API protocolsopenai chat completions
- Catalogued endpoints1
- Websitegroq.com
- Last verified2026-10-02
Observed models & declared pricing
Read from Groq's own catalogue by continuous discovery — declared data, not benchmark results.
| Model | Context | Observed |
|---|
| allam-2-7b | 4k | 2026-10-02 |
| canopylabs/orpheus-arabic-saudi | 4k | 2026-10-02 |
| canopylabs/orpheus-v1-english | 4k | 2026-10-02 |
| meta-llama/llama-prompt-guard-2-22m | 1k | 2026-10-02 |
| meta-llama/llama-prompt-guard-2-86m | 1k | 2026-10-02 |
| openai/gpt-oss-120b | 131k | 2026-10-02 |
| openai/gpt-oss-20b | 131k | 2026-10-02 |
| openai/gpt-oss-safeguard-20b | 131k | 2026-10-02 |
| qwen/qwen3.8-27b | 131k | 2026-10-02 |
| whisper-large-v3 | 0k | 2026-10-02 |
| whisper-large-v3-turbo | 0k | 2026-10-02 |
All providers · How we benchmark