DeepSeekvsPhi

Comparing DeepSeek and Microsoft. Average Podium Score across top models: DeepSeek (71.6) vs Phi (58.4). Flagship showdown: Deepseek V4 Pro High Preview (73.0) vs Phi 3 Medium 4k Instruct (60.0).

DeepSeek (DeepSeek)

High-efficiency, cost-effective reasoning and coding open-weight models.

Avg Score: 71.6
Licensing: Open Weights Available

Phi (Microsoft)

Highly capable Small Language Models (SLMs) from Microsoft Research.

Avg Score: 58.4
Licensing: Open Weights Available

Top Models Showdown

RankModelLabScoreCodingReasoningSpeedPrice / 1M
#1Deepseek V4 Pro High PreviewDeepSeek73.050 t/s$1.1/M
#2Deepseek V4 Flash High PreviewDeepSeek72.050 t/s$1.1/M
#3Deepseek V3.2 Exp ThinkingDeepSeek71.050 t/s$1.1/M
#4Deepseek V3.2 ThinkingDeepSeek71.050 t/s$1.1/M
#5Deepseek V3.2 ExpDeepSeek71.050 t/s$1.1/M
#6Phi 3 Medium 4k InstructMicrosoft60.050 t/s$3/M
#7Wizardlm 70bMicrosoft59.050 t/s$3/M
#8Phi 3 Small 8k InstructMicrosoft59.050 t/s$3/M
#9Wizardlm 13bMicrosoft57.050 t/s$3/M
#10Phi 3 Mini 4k Instruct June 2024Microsoft57.050 t/s$3/M
0 / 4 Models Selected