OpenAIAPI Host BenchmarksEstimated
GPT-6 Astra — API Providers & Speed Matrix
Modelled throughput, TTFT latency and prompt-cache pricing estimates across cloud hosting endpoints, derived from official list prices and public benchmarks.
GPT-6 Astra is a proprietary model served through OpenAI's official API. Third-party hosting is not available — the matrix below covers the official endpoint only. See the OpenAI provider page for details.
Fastest provider
82 t/s
Lowest latency (TTFT)
143 ms P50
Best price
$10 / $50
Blended Pricing Workload Calculator
Switch between standard conversational ratios, agentic caching loops, and heavy RAG workloads.
| Provider & Host | Hardware Stack | Output Speed (TPS) | Latency TTFT (P50) | Input Price ($/1M) | Output Price ($/1M) | Cache Read ($/1M) | Blended Price ($/1M) | SLA |
|---|---|---|---|---|---|---|---|---|
OpenAIOfficial | Official cloud API | 82 t/s | 143 ms (200 ms p90) | $10 | $50 | $1 | $20 | 99.99% |
Figures are modelled estimates based on published list prices and public throughput benchmarks — not measured on live endpoints. Verify current pricing and terms with each host before committing.