Independent analysis of AI inference performance, pricing and infrastructure, based on InferenceBench data.
GPT-OSS 120B spans 179–474 ms median TTFT across measured providers — a 2.6× spread on identical weights.
Support independent AI infrastructure research.