MuseAPI Host BenchmarksEstimated
Muse Spark 1.3 (xhigh) — API Providers & Speed Matrix
Modelled throughput, TTFT latency and prompt-cache pricing estimates across cloud hosting endpoints, derived from official list prices and public benchmarks.
Muse Spark 1.3 (xhigh) is a proprietary model served through Muse's official API. Third-party hosting is not available — the matrix below covers the official endpoint only. See the Muse provider page for details.
Fastest provider
65 t/s
Lowest latency (TTFT)
800 ms P50
Best price
$2.5 / $10
Blended Pricing Workload Calculator
Switch between standard conversational ratios, agentic caching loops, and heavy RAG workloads.
| Provider & Host | Hardware Stack | Output Speed (TPS) | Latency TTFT (P50) | Input Price ($/1M) | Output Price ($/1M) | Cache Read ($/1M) | Blended Price ($/1M) | SLA |
|---|---|---|---|---|---|---|---|---|
MuseOfficial | Official cloud API | 65 t/s | 800 ms (1.12 s p90) | $2.5 | $10 | $0.25 | $4.38 | 99.99% |
Figures are modelled estimates based on published list prices and public throughput benchmarks — not measured on live endpoints. Verify current pricing and terms with each host before committing.