GitHub Copilot最佳AI模型
GitHub Copilot provides AI-powered code suggestions directly in your editor. It uses models for inline completions, chat, and code review. Fast, streaming-capable models work best.
GitHub Copilot的重要因素
Best Models for GitHub Copilot
Top 15 by tool-optimized score
GitHub Copilot全部模型排名(435个模型)
Scored by: benchmark performance (90%) from MMLU, GPQA, HumanEval, SWE-bench, and 15+ standardized evaluations, with capabilities and context as tiebreakers (10%).
| # | 模型 | 评分 | 输出$/百万 |
|---|---|---|---|
| 1 | GLM 5.3 Flash Arena Elo: 1475 | 90 | $0.500 |
| 2 | MiMo-V2.5-Pro Arena Elo: 1467 | 89 | $0.870 |
| 3 | Muse Spark 1.1 Arena Elo: 1493 | 88 | $4.25 |
| 4 | Hy3 Arena Elo: 1456 | 88 | $0.330 |
| 5 | Gemma 4 31B Arena Elo: 1451 | 88 | $0.340 |
| 6 | GPT-5.6 Luna Pro (batch) | 87 | $0.600 |
| 7 | GPT-5.6 Luna (batch) | 87 | $0.600 |
| 8 | MiMo-V2.5 Arena Elo: 1434 | 87 | $0.280 |
| 9 | Gemma 4 26B A4B Arena Elo: 1438 | 87 | $0.300 |
| 10 | Llama 4 Maverick HumanEval: 89.5% | 87 | $0.652 |
| 11 | Kimi K2.6 Arena Elo: 1460 | 86 | $4.00 |
| 12 | GLM 5.1 Arena Elo: 1466 | 86 | $3.04 |
| 13 | GLM 5 Arena Elo: 1458 | 86 | $1.92 |
| 14 | DeepSeek V3.2 Arena Elo: 1425 | 86 | $0.400 |
| 15 | DeepSeek V3.2 Exp Arena Elo: 1422 | 86 | $0.410 |
| 16 | Llama 3.3 70B Instruct HumanEval: 88.4% | 86 | $0.320 |
| 17 | GPT-4o-mini HumanEval: 87.2% | 86 | $0.600 |
| 18 | Gemini 3.5 Flash Lite Arena Elo: 1456 | 85 | $2.50 |
| 19 | Qwen3.7 Plus Arena Elo: 1456 | 85 | $1.28 |
| 20 | Hy3 preview Arena Elo: 1413 | 85 | $0.600 |
| 21 | Gemma 4 31B (free) | 85 | Free |
| 22 | DeepSeek V3.1 Arena Elo: 1417 | 85 | $0.950 |
| 23 | Qwen3.8 27B Arena Elo: 1437 | 84 | $3.00 |
| 24 | Inkling Arena Elo: 1440 | 84 | $4.05 |
| 25 | GPT-5.6 Luna Pro | 84 | $1.20 |
| 26 | GPT-5.6 Luna | 84 | $1.20 |
| 27 | MiniMax M3 Arena Elo: 1441 | 84 | $1.20 |
| 28 | Gemini 3.5 Flash Arena Elo: 1477 | 84 | $9.00 |
| 29 | Claude Opus 4.7 (batch) | 84 | $12.50 |
| 30 | Qwen3.6 Plus Arena Elo: 1443 | 84 | $1.95 |
| 31 | Qwen3.5 397B A17B Arena Elo: 1442 | 84 | $3.50 |
| 32 | GLM 4.7 Arena Elo: 1442 | 84 | $1.75 |
| 33 | GPT-5.2 Chat Arena Elo: 1476 | 84 | $14.00 |
| 34 | Grok 4.7 | 83 | $4.80 |
| 35 | Grok 4.5 Arena Elo: 1468 | 83 | $6.00 |
| 36 | Claude Opus 4.8 (batch) | 83 | $12.50 |
| 37 | Grok 4.3 | 83 | $2.50 |
| 38 | Qwen3.6 Max Preview Arena Elo: 1460 | 83 | $6.16 |
| 39 | GLM 5V Turbo Arena Elo: 1433 | 83 | $4.00 |
| 40 | Grok 4.20 | 83 | $2.50 |
| 41 | Gemini 3.1 Flash Lite Preview Arena Elo: 1432 | 83 | $1.50 |
| 42 | Qwen3.5-Flash Arena Elo: 1397 | 83 | $0.260 |
| 43 | Step 3.5 Flash Arena Elo: 1394 | 83 | $0.300 |
| 44 | Gemini 3 Flash Preview (batch) | 83 | $1.50 |
| 45 | GPT-5.1-Codex-Mini | 83 | $2.00 |
| 46 | GLM 4.6 Arena Elo: 1425 | 83 | $1.75 |
| 47 | Gemma 4 26B A4B (free) | 82 | Free |
| 48 | MiniMax M2.7 Arena Elo: 1415 | 82 | $1.20 |
| 49 | GPT-5.4 Nano (batch) | 82 | $0.625 |
| 50 | GPT-5.4 (batch) | 82 | $7.50 |
Based on our analysis of coding benchmarks, capability matching, and pricing, GLM 5.3 Flash currently ranks #1 for GitHub Copilot. Rankings are rebuilt as benchmark, pricing, and provider data refresh.
We score models using benchmark performance (90%) from LMArena, HumanEval, SWE-bench, MMLU, and 15+ standardized evaluations. Capabilities and context serve as tiebreakers (10%). Only models with the capabilities GitHub Copilot needs are included in the tool-specific rankings.
We currently track 435 AI models compatible with GitHub Copilot. This includes models from OpenAI, Anthropic, Google, DeepSeek, and other providers accessible via API.
Many open-source models are compatible with GitHub Copilot through API providers like OpenRouter, Together AI, and Groq. Check our rankings to see which open-source models perform best.
Rankings refresh whenever the underlying benchmark, pricing, and catalog sources refresh. That means some signals update faster than others, and the page reflects the latest verified source data available.