OpenAIAPI Host BenchmarksEstimated

GPT-6 Astra — API Providers & Speed Matrix

Modelled throughput, TTFT latency and prompt-cache pricing estimates across cloud hosting endpoints, derived from official list prices and public benchmarks.

GPT-6 Astra is a proprietary model served through OpenAI's official API. Third-party hosting is not available — the matrix below covers the official endpoint only. See the OpenAI provider page for details.
Fastest provider
82 t/s
OpenAI (Official cloud API)
Lowest latency (TTFT)
143 ms P50
OpenAI (TTFT P90: 0.2s)
Best price
$10 / $50
OpenAI (In / Out $/1M)

Blended Pricing Workload Calculator

Switch between standard conversational ratios, agentic caching loops, and heavy RAG workloads.

Provider & HostHardware StackOutput Speed (TPS)Latency TTFT (P50)Input Price ($/1M)Output Price ($/1M)Cache Read ($/1M)Blended Price ($/1M)SLA
OpenAIOfficial
Official cloud API82 t/s143 ms (200 ms p90)$10$50$1$2099.99%

Figures are modelled estimates based on published list prices and public throughput benchmarks — not measured on live endpoints. Verify current pricing and terms with each host before committing.

0 / 4
Сравнить →