OpenAI

o3รุ่นเก่า

Chain-of-thought reasoning model, a former math and code benchmark leader.

#429เปิดตัวเมื่อ 15 เม.ย. 2568·Proprietary·Not disclosed
textimage200K ctxreasoning
API providersตารางอันดับ
Intelligence
32.0 #429
Top 61% tier
Output speed
90 t/s
TTFT ~0.8s · ITL 11.1 ms/tok
Blended price (3:1)
$4.375 /1M
In $2.5/M · Out $10/M
Context
200K
≈ 500 A4 pages · max out 16384

Quality Domain Radar

6-Axis Frontier Evaluation vs Category Average & Flagship

o3
Claude Mythos Preview (#1)
Average (68)
Coding (35)Mathematics (37)Hard Science (GPQA) (38)Agentic & Tool Use (30)Knowledge (MMLU) (32)Vision / Multimodal (51.6)

Radar axes are approximated from category scores where direct benchmark measurements are missing — see ระเบียบวิธี.

API deployment

Compare hosts & latency

Throughput, TTFT and cache pricing across Groq, Cerebras, Together, DeepInfra, Fireworks, Bedrock and Azure.

GroqCerebrasTogether AIDeepInfraFireworks AIAWS BedrockAzure AI
View providers matrix →

คะแนนแยกตามแหล่งข้อมูล

คะแนนแยกตามหมวดหมู่

ผลการทดสอบ Benchmark

รายละเอียดโมเดล

วันที่เปิดตัว
15 เม.ย. 2568
พารามิเตอร์
Not disclosed
สัญญาอนุญาต
Proprietary
ขนาดคอนเทกซ์
200,000 tokens
ราคาขาเข้า
$2.5/M
ราคาขาออก
$10/M
ความเร็วขาออก
90 t/s
เวลาถึงโทเค็นแรก

Estimated monthly bill

Estimated monthly spend
$0.00
Calculating savings vs GPT-4o...

Cite & embed

GitHub Markdown badge
[![LLMPodium Rank](https://llmpodium.com/badge/o3/rank.svg)](https://llmpodium.com/models/o3)
BibTeX
@misc{llmpodium2026o3,
  title = {Benchmark Analysis and Intelligence Evaluation of o3},
  author = {LLMPodium Research},
  year = {2026},
  url = {https://llmpodium.com/models/o3}
}

More from OpenAI

โมเดลทางเลือก

ทางเลือกที่ใกล้เคียงที่สุดสำหรับ o3 โดยอิงตามความสามารถ ราคา และคะแนนเบนช์มาร์ก

0 / 4