RANKING / 第三方榜单
模型榜单
数据来源:Artificial Analysis · 通道 page_parse · 批次时间 · 共 656 个模型 · 其中 49 条同时有分值与价格(可进斩杀线图)
BY DIMENSION / 按维度
知识与推理 榜
知识准确性与难题(omniscience / gpqa / hle) · 价格为美元 / 百万 tokens(第三方原始口径)
| # | 模型 | 知识与推理 (知识准确性与难题(omniscience / gpqa / hle)) | 综合分 | 较上批 | 输入价 | 输出价 | 吞吐 | |
|---|---|---|---|---|---|---|---|---|
| 172 | o3-pro OpenAI 本批新增 推理档位 | 84.5 | 21.87 | — | ¥134.2 | ¥536.82 | — | 200K |
| 327 | o1-preview OpenAI 本批新增 推理档位 | 76.5 | 11.38 | — | ¥110.72 | ¥442.87 | — | 128K |
| 1 | Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback) Anthropic · 档案:Claude Fable 5.1 本批新增 推理档位 | 70.5 | 53.35 | — | ¥67.1 | ¥335.51 | 65.8181 | 1000K |
| 2 | Claude Fable 5.1 (Adaptive Reasoning, Xhigh Effort, Default Fallback) Anthropic · 档案:Claude Fable 5.1 本批新增 推理档位 | 69.8 | 53.20 | — | ¥67.1 | ¥335.51 | 61.8045 | 1000K |
| 7 | Claude Opus 5 (Adaptive Reasoning, Max Effort) Anthropic · 档案:Claude Opus 5 本批新增 推理档位 | 67.3 | 50.78 | — | ¥33.55 | ¥167.76 | 56.4472 | 1000K |
| 9 | Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) Anthropic · 档案:Claude Fable 5 本批新增 推理档位 | 67.0 | 49.63 | — | ¥67.1 | ¥335.51 | 67.872 | 1000K |
| 5 | Claude Fable 5.1 (Adaptive Reasoning, High Effort, Default Fallback) Anthropic · 档案:Claude Fable 5.1 本批新增 推理档位 | 66.8 | 51.15 | — | ¥67.1 | ¥335.51 | 55.7654 | 1000K |
| 8 | Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) Anthropic · 档案:Claude Opus 5 本批新增 推理档位 | 66.6 | 49.68 | — | ¥33.55 | ¥167.76 | 54.4 | 1000K |
| 3 | GPT-6 Astra (max) OpenAI · 档案:GPT-6 Astra 本批新增 推理档位 | 66.4 | 52.67 | — | ¥67.1 | ¥335.51 | 65.2837 | 1000K |
| 4 | GPT-6 Astra (xhigh) OpenAI · 档案:GPT-6 Astra 本批新增 推理档位 | 65.9 | 52.39 | — | ¥67.1 | ¥335.51 | 54.8871 | 1000K |
| 12 | Claude Opus 5 (Adaptive Reasoning, High Effort) Anthropic · 档案:Claude Opus 5 本批新增 推理档位 | 64.9 | 48.12 | — | ¥33.55 | ¥167.76 | 57.3799 | 1000K |
| 6 | GPT-6 Astra (high) OpenAI · 档案:GPT-6 Astra 本批新增 推理档位 | 64.6 | 50.92 | — | ¥67.1 | ¥335.51 | 54.4358 | 1000K |
| 14 | GPT-5.6 Sol (max) OpenAI · 档案:GPT-5.6 Sol 本批新增 推理档位 | 64.4 | 46.97 | — | ¥26.84 | ¥134.2 | 84.2378 | 1000K |
| 11 | Claude Fable 5.1 (Adaptive Reasoning, Medium Effort, Default Fallback) Anthropic · 档案:Claude Fable 5.1 本批新增 推理档位 | 64.3 | 48.92 | — | ¥67.1 | ¥335.51 | 56.5061 | 1000K |
| 10 | GPT-6 Astra (medium) OpenAI · 档案:GPT-6 Astra 本批新增 推理档位 | 63.9 | 49.57 | — | ¥67.1 | ¥335.51 | 53.6062 | 1000K |
| 26 | GPT-5.6 Sol (xhigh) OpenAI · 档案:GPT-5.6 Sol 本批新增 推理档位 | 62.9 | 44.01 | — | ¥26.84 | ¥134.2 | 72.3042 | 1000K |
| 408 | Sonar Reasoning Perplexity · 未匹配本地档案 本批新增 推理档位 | 62.3 | 8.67 | — | — | — | — | 127K |
| 74 | GPT-5.3 Codex (xhigh) OpenAI · 档案:GPT-5.3 Codex 本批新增 推理档位 | 62.3 | 32.50 | — | ¥11.74 | ¥93.94 | 175.6169 | 400K |
| 22 | Claude Opus 5 (Adaptive Reasoning, Medium Effort) Anthropic · 档案:Claude Opus 5 本批新增 推理档位 | 62.3 | 44.83 | — | ¥33.55 | ¥167.76 | 57.7063 | 1000K |
| 100 | Gemini 3 Pro Preview (high) Google · 档案:Gemini 3 Pro Preview 本批新增 推理档位 | 62.1 | 27.96 | — | ¥13.42 | ¥80.52 | — | 1000K |
| 30 | GPT-5.6 Sol (high) OpenAI · 档案:GPT-5.6 Sol 本批新增 推理档位 | 61.5 | 42.35 | — | ¥26.84 | ¥134.2 | 71.2031 | 1000K |
| 69 | Gemini 3.5 Flash (medium) Google · 档案:Gemini 3.5 Flash 本批新增 推理档位 | 61.5 | 33.63 | — | ¥10.07 | ¥60.39 | 240.4431 | 1000K |
| 19 | GPT-6 Astra (low) OpenAI · 档案:GPT-6 Astra 本批新增 推理档位 | 61.3 | 45.78 | — | ¥67.1 | ¥335.51 | 53.3243 | 1000K |
| 15 | Claude Fable 5.1 (Adaptive Reasoning, Low Effort, Default Fallback) Anthropic · 档案:Claude Fable 5.1 本批新增 推理档位 | 61.2 | 46.82 | — | ¥67.1 | ¥335.51 | 55.2372 | 1000K |
| 13 | Muse Spark 1.3 (max) Meta · 档案:Muse Spark 1.3 本批新增 推理档位 | 61.1 | 48.09 | — | ¥8.39 | ¥28.52 | 222.1289 | 1000K |
| 34 | Gemini 3.8 Flash (high) Google · 档案:Gemini 3.8 Flash 本批新增 推理档位 | 60.8 | 40.93 | — | ¥5.03 | ¥25.16 | 343.5202 | 1000K |
| 24 | Grok 4.6 (high) xAI · 档案:Grok 4.6 本批新增 推理档位 | 60.3 | 44.31 | — | ¥13.42 | ¥40.26 | 60.4124 | 500K |
| 45 | Gemini 3.7 Flash (high) Google · 档案:Gemini 3.7 Flash 本批新增 推理档位 | 60.3 | 39.06 | — | ¥5.03 | ¥25.16 | 334.6745 | 1000K |
| 115 | Gemini 3 Flash Preview (Reasoning) Google · 档案:Gemini 3 Flash Preview 本批新增 推理档位 | 59.9 | 26.33 | — | ¥3.36 | ¥20.13 | 218.4829 | 1000K |
| 21 | Muse Spark 1.3 (xhigh) Meta · 档案:Muse Spark 1.3 本批新增 推理档位 | 59.9 | 45.07 | — | ¥8.39 | ¥28.52 | 260.3535 | 1000K |
| 28 | Kimi K3 (max) Moonshot AI · 档案:Kimi K3 本批新增 推理档位 | 59.8 | 43.59 | — | ¥20.13 | ¥100.65 | 40.6343 | 1049K |
| 48 | GPT-5.5 (xhigh) OpenAI · 档案:GPT-5.5 本批新增 推理档位 | 59.8 | 38.36 | — | ¥33.55 | ¥201.31 | 110.1435 | 922K |
| 44 | GPT-5.6 Sol (medium) OpenAI · 档案:GPT-5.6 Sol 本批新增 推理档位 | 59.4 | 39.24 | — | ¥26.84 | ¥134.2 | 63.6866 | 1000K |
| 25 | Grok 4.6 (xhigh) xAI · 档案:Grok 4.6 本批新增 推理档位 | 59.3 | 44.20 | — | ¥13.42 | ¥40.26 | 58.1876 | 500K |
| 33 | Claude Opus 4.8 (Adaptive Reasoning, Max Effort) Anthropic · 档案:Claude Opus 4.8 本批新增 推理档位 | 59.1 | 41.79 | — | ¥33.55 | ¥167.76 | 56.6546 | 1000K |
| 52 | GPT-5.5 (high) OpenAI · 档案:GPT-5.5 本批新增 推理档位 | 58.9 | 36.98 | — | ¥33.55 | ¥201.31 | 97.7378 | 922K |
| 77 | Claude Opus 4.6 (Adaptive Reasoning, Max Effort) Anthropic · 档案:Claude Opus 4.6 本批新增 推理档位 | 58.8 | 31.95 | — | ¥33.55 | ¥167.76 | 42.0691 | 1000K |
| 39 | Gemini 3.8 Flash (medium) Google · 档案:Gemini 3.8 Flash 本批新增 推理档位 | 58.5 | 39.77 | — | ¥5.03 | ¥25.16 | — | 1000K |
| 29 | Grok 4.6 (medium) xAI · 档案:Grok 4.6 本批新增 推理档位 | 58.2 | 42.84 | — | ¥13.42 | ¥40.26 | 59.7579 | 500K |
| 47 | Grok 4.5 (high) xAI · 档案:Grok 4.5 本批新增 推理档位 | 57.7 | 38.81 | — | ¥13.42 | ¥40.26 | 58.093 | 500K |
| 41 | Muse Spark 1.2 (xhigh) Meta · 档案:Muse Spark 1.2 本批新增 推理档位 | 57.6 | 39.58 | — | ¥8.39 | ¥28.52 | 186.7485 | 1049K |
| 54 | DeepSeek V4 Pro 0813 (Reasoning, Max Effort) DeepSeek · 档案:DeepSeek V4 Pro 0813 本批新增 推理档位 | 57.5 | 36.00 | — | ¥8.86 | ¥26.57 | 73.73 | 1000K |
| 84 | GPT-5.2 (xhigh) OpenAI · 档案:GPT-5.2 本批新增 推理档位 | 57.4 | 30.45 | — | ¥11.74 | ¥93.94 | 85.3266 | 400K |
| 31 | GPT-5.6 Terra (max) OpenAI · 档案:GPT-5.6 Terra 本批新增 推理档位 | 57.2 | 42.08 | — | ¥13.42 | ¥80.52 | 104.6262 | 1000K |
| 43 | Claude Opus 5 (Adaptive Reasoning, Low Effort) Anthropic · 档案:Claude Opus 5 本批新增 推理档位 | 57.0 | 39.35 | — | ¥33.55 | ¥167.76 | 59.9228 | 1000K |
| 65 | GPT-5.5 (medium) OpenAI · 档案:GPT-5.5 本批新增 推理档位 | 56.9 | 33.80 | — | ¥33.55 | ¥201.31 | 95.2277 | 922K |
| 40 | Gemini 3.7 Flash (medium) Google · 档案:Gemini 3.7 Flash 本批新增 推理档位 | 56.8 | 39.62 | — | ¥5.03 | ¥25.16 | 317.8079 | 1000K |
| 20 | Qwen3.8 Max (0902) Alibaba · 档案:Qwen3.8 Max 本批新增 推理档位 | 56.5 | 45.42 | — | ¥13.42 | ¥40.26 | 39.2448 | 984K |
| 71 | GPT-5.6 Sol (low) OpenAI · 档案:GPT-5.6 Sol 本批新增 推理档位 | 56.5 | 33.47 | — | ¥26.84 | ¥134.2 | 67.0664 | 1000K |
| 23 | GLM-5.3 (max) Zhipu AI · 档案:GLM-5.3 本批新增 推理档位 | 56.3 | 44.78 | — | ¥9.39 | ¥29.52 | 61.0413 | 1000K |
币种口径:价格已折算为人民币(1 USD = 6.710202 CNY,来源 open.er-api.com · 2026-09-22);原始口径为第三方币种。
同一模型的多个推理档位(max / xhigh / high…)在榜单里是多行;模型档案把它们折叠成一个条目
(取分最高的档位),点进去可以看到该模型的汇总读数。
想比较性价比请看 斩杀线坐标图 ·
本页只展示已验证有效的批次(校验规则见方法学)。
维度口径:知识与推理 —— 知识准确性与难题(omniscience / gpqa / hle);
复合维度(编程/智能体/知识)由该组内的评测项**取均值**得到(缺项不参与,不按 0 计),
原始维度值可在模型档案页查看。所有维度都来自同一批快照,切换维度不会重新抓取。