Models
Coding Models
面向代码生成与软件开发的 AI 模型。收录 12 个模型,按 CODX Score 排序比较。
| # | 模型 | 开发公司 | Context | Coding | Agent | API | Open Source | CODX Score |
|---|---|---|---|---|---|---|---|---|
| 01 | GPT-5-Codex | OpenAI | 1M tokens 级别上下文 | 97 | 95 | ● | — | 95 |
| 02 | Claude Opus | Anthropic | 200K tokens 上下文窗口 | 96 | 94 | ● | — | 93 |
| 03 | Claude Sonnet | Anthropic | 200K tokens 上下文窗口 | 94 | 92 | ● | — | 92 |
| 04 | Gemini 3 Pro | 1M+ tokens 上下文窗口 | 92 | 90 | ● | — | 90 | |
| 05 | DeepSeek-V3 | DeepSeek | 128K tokens 上下文窗口 | 91 | 88 | ● | ✓ | 89 |
| 06 | Qwen3-Coder | Alibaba | 256K tokens 上下文窗口 | 90 | 88 | ● | ✓ | 88 |
| 07 | GPT-5.1 | OpenAI | 400K tokens 上下文窗口 | 89 | 86 | ● | — | 88 |
| 08 | Gemini 2.5 Pro | 1M tokens 上下文窗口 | 88 | 86 | ● | — | 87 | |
| 09 | DeepSeek-R1 | DeepSeek | 128K tokens 上下文窗口 | 88 | 84 | ● | ✓ | 86 |
| 10 | Kimi k2 | Moonshot AI | 256K tokens 上下文窗口 | 87 | 88 | ● | ✓ | 86 |
| 11 | GLM-4.6 | Zhipu AI (智谱) | 200K tokens 上下文窗口 | 84 | 82 | ● | ✓ | 83 |
| 12 | Codestral | Mistral AI | 256K tokens 上下文窗口 | 84 | 80 | ● | ✓ | 83 |
01
GPT-5-Codex
OpenAI
95
Excellent
Context1M tokens 级别上下文Coding97Agent95协议API
02
Claude Opus
Anthropic
93
Excellent
Context200K tokens 上下文窗口Coding96Agent94协议API
03
Claude Sonnet
Anthropic
92
Excellent
Context200K tokens 上下文窗口Coding94Agent92协议API
04
Gemini 3 Pro
Google
90
Excellent
Context1M+ tokens 上下文窗口Coding92Agent90协议API
05
DeepSeek-V3
DeepSeek
89
Great
Context128K tokens 上下文窗口Coding91Agent88协议API / 开源
06
Qwen3-Coder
Alibaba
88
Great
Context256K tokens 上下文窗口Coding90Agent88协议API / 开源
07
GPT-5.1
OpenAI
88
Great
Context400K tokens 上下文窗口Coding89Agent86协议API
08
Gemini 2.5 Pro
Google
87
Great
Context1M tokens 上下文窗口Coding88Agent86协议API
09
DeepSeek-R1
DeepSeek
86
Great
Context128K tokens 上下文窗口Coding88Agent84协议API / 开源
10
Kimi k2
Moonshot AI
86
Great
Context256K tokens 上下文窗口Coding87Agent88协议API / 开源
11
GLM-4.6
Zhipu AI (智谱)
83
Great
Context200K tokens 上下文窗口Coding84Agent82协议API / 开源
12
Codestral
Mistral AI
83
Great
Context256K tokens 上下文窗口Coding84Agent80协议API / 开源
Coding / Agent 分数为 CODX 基于公开资料与实测的编辑评分,满分 100。完整方法见 Methodology。