Skip to main content
LLMPodium
순위표
🏆 Overall LeaderboardComposite index across all flagship LLMs💻 Coding & SWE-BenchReal-world programming & software agent benchmark🧠 Deep ReasoningGPQA Diamond, AIME & complex logic🤖 Autonomous AgentsOSWorld, Toolathlon & multi-turn execution🔓 Open WeightsDeepSeek, Qwen, Llama & Mistral
아레나
⚔️ Human Arena BattlesBlind pairwise Elo rankings (700+ models)⚖️ Side-by-Side CompareHead-to-head metric comparison of up to 4 models💰 Cost-per-Task CalculatorEstimate API economics across real workflows🎯 Model RecommenderFind the optimal model for your budget and speed
모델 목록
📦 Model DirectoryDetailed profiles, context windows & pricing🏢 AI Labs & ProvidersOpenAI, Anthropic, Google, DeepSeek, Meta📊 Benchmark MatrixEvaluation methodologies and leader tables
nav.resources
📐 Podium MethodologyMathematical aggregation formula explained📰 Changelog & NewsDaily model updates and new evaluations✍️ Research & ArticlesIn-depth AI benchmarking whitepapers❓ Frequently Asked QuestionsCommon questions on scores and ranking📖 LLM GlossaryDefinitions of TTFT, TPS, Elo, MoE & tokens
🇺🇸 EnglishEN🇨🇳 中文ZH🇮🇳 हिन्दीHI🇪🇸 EspañolES🇫🇷 FrançaisFR🇸🇦 العربيةAR🇧🇩 বাংলাBN🇧🇷 PortuguêsPT🇷🇺 РусскийRU🇵🇰 اردوUR🇮🇩 Bahasa IndonesiaID🇩🇪 DeutschDE🇯🇵 日本語JA🇰🇷 한국어KO🇹🇭 ไทยTH🇮🇹 ItalianoIT
순위표
🏆 순위표
Overall Composite LeaderboardCoding & Software AgentsDeep Reasoning & MathAutonomous Agent SwarmsOpen Weights & Self-Hosted
⚔️ 아레나 & Compare
Human Battle Arena (700+ Models)Side-by-Side Model CompareCost-per-Task CalculatorInteractive Model Finder
📦 Directories & Evals
Model Specifications DirectoryAI Labs & Cloud ProvidersBenchmark Methodologies
📖 Research & Info
Scoring MethodologyResearch Blog & ArticlesChangelog & New ModelsFrequently Asked QuestionsTechnical LLM GlossaryAbout LLMPodium
🌐 Language / Язык / 语言
🇺🇸English🇨🇳中文🇮🇳हिन्दी🇪🇸Español🇫🇷Français🇸🇦العربية🇧🇩বাংলা🇧🇷Português🇷🇺Русский🇵🇰اردو🇮🇩Bahasa Indonesia🇩🇪Deutsch🇯🇵日本語🇰🇷한국어🇹🇭ไทย🇮🇹Italiano
Explore Full Leaderboard →
  1. 홈
  2. /최고 모델

🏆 최고 모델

랭킹은 독립적인 벤치마크 평가를 기반으로 합니다.

💻

코딩 최고의 AI 모델

코딩 성능 기반 언어 모델 순위.

🧠

추론 최고의 AI 모델

추론 능력 기반 언어 모델 순위.

📚

지식 최고의 AI 모델

지식 정확도 기반 언어 모델 순위.

👁️

비전 최고의 AI 모델

시각 이해 기반 멀티모달 모델 순위.

💰

최고 가성비 AI 모델

달러당 지능 기반 모델 순위.

⚡

가장 빠른 AI 모델

생성 속도 기반 모델 순위.

🏆

가장 지능적인 AI 모델

종합 지능 기반 모델 순위.

🔢

수학 최고의 AI 모델

수학적 추론 기반 모델 순위.

📄

긴 컨텍스트 최고의 AI 모델

긴 컨텍스트 능력 기반 모델 순위.

🤖

에이전트 최고의 AI 모델

에이전트 성능 기반 모델 순위.

🛡️

가장 안전한 AI 모델

안전성과 정렬 기반 모델 순위.

📋

지시 따르기 최고의 AI 모델

지시 따르기 능력 기반 모델 순위.

LLMPodium

The definitive open LLM benchmark aggregator and leaderboard. Continuous evaluations across 719+ models and 25+ benchmarks.

Leaderboards synced daily
순위표
  • 순위표
  • 💻 코딩
  • 🧠 추론
  • 📐 수학
  • 🤖 에이전트
  • 🔓 오픈 웨이트
  • ⚡ 속도
  • 💰 가성비
도구
  • 아레나 (719+ models)
  • 비교 Side-by-Side
  • 벤치마크 Matrix
  • 추천
  • 모델 목록 Directory
  • AI 제공업체
  • Cost per Task Calculator
  • #1 Ranked Model Profile
리소스
  • 뉴스 Feed
  • 블로그 & Research
  • Best AI Models 2026
  • Enterprise AI Use Cases
  • 평가 방법
  • 용어집
  • FAQ
  • 소개
© 2026 LLMPodium. All rights reserved.·데이터는 공개 벤치마크에서 정기적으로 업데이트됩니다.
llms.txtRSS FeedOpen API