ClaudevsPhi

Comparing Anthropic and Microsoft. Average Podium Score across top models: Claude (84.6) vs Phi (58.4). Flagship showdown: Claude Mythos Preview (97.4) vs Phi 3 Medium 4k Instruct (60.0).

Claude (Anthropic)

Industry-leading frontier reasoning and coding models from Anthropic.

Avg Score: 84.6
Licensing: Proprietary API

Phi (Microsoft)

Highly capable Small Language Models (SLMs) from Microsoft Research.

Avg Score: 58.4
Licensing: Open Weights Available

Top Models Showdown

RankModelLabScoreCodingReasoningSpeedPrice / 1M
#1Claude Mythos PreviewAnthropic97.480.998.880 t/s$15/M
#2Claude Fable 5Anthropic93.484.181.071 t/s$50/M
#3Claude Opus 5Anthropic81.470.482.360 t/s$25/M
#4Claude Opus 4.8Anthropic76.058.082.850 t/s$15/M
#5Claude Opus 4 6 ThinkingAnthropic75.050 t/s$15/M
#6Phi 3 Medium 4k InstructMicrosoft60.050 t/s$3/M
#7Wizardlm 70bMicrosoft59.050 t/s$3/M
#8Phi 3 Small 8k InstructMicrosoft59.050 t/s$3/M
#9Wizardlm 13bMicrosoft57.050 t/s$3/M
#10Phi 3 Mini 4k Instruct June 2024Microsoft57.050 t/s$3/M
0 / 4 Models Selected