OpenAI
o3旧版
Chain-of-thought reasoning model, a former math and code benchmark leader.
Intelligence
32.0 #429
Top 61% tier
Output speed
90 t/s
TTFT ~0.8s · ITL 11.1 ms/tok
Blended price (3:1)
$4.375 /1M
In $2.5/M · Out $10/M
Context
200K
≈ 500 A4 pages · max out 16384
Quality Domain Radar
6-Axis Frontier Evaluation vs Category Average & Flagship
o3
Claude Mythos Preview (#1)
Average (68)
Radar axes are approximated from category scores where direct benchmark measurements are missing — see 评分方法.
API deployment
Compare hosts & latency
Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.
GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
各榜单得分
分类能力得分
基准测试分项
模型详细参数
发布日期
2025年4月15日
参数量
Not disclosed
开源协议
Proprietary
上下文窗口
200,000 tokens
输入价格
$2.5/M
输出价格
$10/M
输出速度
90 t/s
首Token延迟
—
Estimated monthly bill
Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...
Cite & embed
GitHub Markdown badge
[](https://llmpodium.com/models/o3)BibTeX
@misc{llmpodium2026o3,
title = {Benchmark Analysis and Intelligence Evaluation of o3},
author = {LLMPodium Research},
year = {2026},
url = {https://llmpodium.com/models/o3}
}More from OpenAI
替代模型推荐
基于能力、定价和基准测试得分与 o3 最接近的替代模型。