DataLearner 标志

Artificial Analysis Intelligence Index AI模型智能指数排行榜

Artificial Analysis Intelligence Index v4.0 综合了10项权威评测基准(GDPval-AA、Terminal-Bench、GPQA Diamond、SciCode等),从数学、科学、编程、推理等多维度对AI模型进行全面评估和排名。

榜首模型

Claude Opus 5.5 (max with fallback)

最高得分

58

模型数量

263

数据版本

2026年10月01日

数据来源: Artificial Analysis

榜单历史快照月份:

排名总表

排名模型名称智能指数机构
AnthropicClaude Opus 5.5 (max with fallback)Anthropic58Anthropic
AnthropicClaude Opus 5.5 (xhigh with fallback)Anthropic56Anthropic
AnthropicClaude Sonnet 5.5 (max with fallback)Anthropic56Anthropic
4AnthropicClaude Opus 5.5 (high with fallback)Anthropic54Anthropic
5AnthropicClaude Fable 5.1Anthropic53Anthropic
6AnthropicClaude Fable 5.1Anthropic53Anthropic
7OpenAIGPT-6 Astra (max)OpenAI53OpenAI
8GoogleGemini 4 Argon (high)Google53Google
9OpenAIGPT-6 Astra (xhigh)OpenAI52OpenAI
10AnthropicClaude Sonnet 5.5 (xhigh with fallback)Anthropic52Anthropic
11OpenAIGPT-6.1 Sol (max)OpenAI52OpenAI
12AnthropicClaude Opus 5.5 (medium with fallback)Anthropic51Anthropic
13AnthropicClaude Fable 5.1Anthropic51Anthropic
14OpenAIGPT-6.1 Sol (xhigh)OpenAI51OpenAI
15OpenAIGPT-6 Astra (high)OpenAI51OpenAI
16OpenAIGPT-6.1 Sol (high)OpenAI50OpenAI
17OpenAIGPT-6 Astra (medium)OpenAI50OpenAI
18AnthropicClaude Fable 5.1Anthropic49Anthropic
19MetaMuse Spark 1.3 (max)Meta48Meta
20OpenAIGPT-6.1 Sol (medium)OpenAI48OpenAI
21OpenAIGPT-6 Sol (max)OpenAI48OpenAI
22AnthropicClaude Fable 5.1Anthropic47Anthropic
23AnthropicClaude Sonnet 5.5 (high with fallback)Anthropic47Anthropic
24SpaceXAIGrok 4.7 (xhigh)SpaceXAI46SpaceXAI
25SpaceXAIGrok 4.7 (high)SpaceXAI46SpaceXAI
26MiMo-V2.6-ProXiaomi46Xiaomi
27OpenAIGPT-6 Astra (low)OpenAI46OpenAI
28AlibabaQwen3.8 Max (0902)Alibaba45Alibaba
29MetaMuse Spark 1.3 (xhigh)Meta45Meta
30智谱AIGLM-5.3 (max)智谱AI45智谱AI
31OpenAIGPT-6 Sol (xhigh)OpenAI44OpenAI
32StepFunStep 5 PreviewStepFun44StepFun
33Moonshot AIKimi K3 (max)Moonshot AI44Moonshot AI
34OpenAIGPT-6 Sol (high)OpenAI43OpenAI
35AnthropicClaude Opus 5.5 (low with fallback)Anthropic42Anthropic
36OpenAIGPT-6.1 Sol (low)OpenAI42OpenAI
37OpenAIGPT-5.6 Terra (max)OpenAI42OpenAI
38Z AIGLM-5.3-FlashZ AI42Z AI
39GoogleGemini 3.8 Flash (high)Google41Google
40AnthropicClaude Sonnet 5.5 (medium with fallback)Anthropic41Anthropic
41AlibabaQwen3.8 2.4T A95BAlibaba40Alibaba
42AlibabaQwen3.8-Flash-NextAlibaba40Alibaba
43OpenAIGPT-6 Sol (medium)OpenAI40OpenAI
44GoogleGemini 3.8 Flash (medium)Google40Google
45DeepSeekDeepSeek V4.1 Flash (max)DeepSeek39DeepSeek
46OpenAIGPT-5.6 Terra (xhigh)OpenAI38OpenAI
47MiMo-V2.6-FlashXiaomi38Xiaomi
48OpenAIGPT-6 Luna (max)OpenAI37OpenAI
49DeepSeek-AIDeepSeek-V4-Pro (max)DeepSeek-AI36DeepSeek-AI
50DeepSeekDeepSeek V4 Flash Vision (max)DeepSeek35DeepSeek
51Z AIGLM-5.3 (low)Z AI34Z AI
52OpenAIGPT-5.6 Terra (high)OpenAI34OpenAI
53JT-4.1 Flash 236B A21BChina Mobile34China Mobile
54OpenAIGPT-6 Sol (low)OpenAI34OpenAI
55OpenAIGPT-6 Luna (xhigh)OpenAI34OpenAI
56阿里巴巴Qwen3.8-27B (xhigh)阿里巴巴34阿里巴巴
57Motif 3Motif Technologies34Motif Technologies
58GoogleGemini 3.8 Flash (low)Google33Google
59OpenAIGPT-5.3 Codex (xhigh)OpenAI33OpenAI
60OpenAIGPT-6 Luna (high)OpenAI32OpenAI
61K2 Horizon 375B A23BMBZUAI Institute of Foundation Models31MBZUAI Institute of Foundation Models
62OpenAIGPT-5.6 Terra (medium)OpenAI30OpenAI
63Moonshot AIKimi K3 (low)Moonshot AI30Moonshot AI
64Google DeepMindGemini 3.1 Pro PreviewGoogle DeepMind30Google DeepMind
65OpenAIGPT-6 Luna (medium)OpenAI29OpenAI
66MiniMaxAIMiniMax M3MiniMaxAI29MiniMaxAI
67Nex-N2-ProNex AGI28Nex AGI
68Solar Pro 4Upstage28Upstage
69OpenAIGPT-6 Sol (non-reasoning)OpenAI28OpenAI
70阿里巴巴Qwen3.8-27B (medium)阿里巴巴28阿里巴巴
71OpenAIGPT-5.6 Terra (low)OpenAI27OpenAI
72JT-4.1 Flash 236B A21B (non-reasoning)China Mobile27China Mobile
73Quasar 438B (max)Multiverse Computing27Multiverse Computing
74Apodex 1.1Apodex26Apodex
75阿里巴巴Qwen3.8-27B (low)阿里巴巴26阿里巴巴
76OpenAIGPT-5.5 InstantOpenAI26OpenAI
77Moonshot AIKimi K2.7 CodeMoonshot AI26Moonshot AI
78Inkling SmallThinking Machines26Thinking Machines
79K2 Horizon MoVA 36B A4BInstitute of Foundation Models25Institute of Foundation Models
80腾讯AI实验室Hy3腾讯AI实验室25腾讯AI实验室
81阿里巴巴Qwen3.7-Plus阿里巴巴25阿里巴巴
82Inkling (xhigh)Thinking Machines25Thinking Machines
83Solar Open2 250BUpstage25Upstage
84DeepSeekDeepSeek V4.1 Flash (non-reasoning)DeepSeek25DeepSeek
85Ling-3.0-flash-VLInclusionAI25InclusionAI
86Solar Mini 4Upstage24Upstage
87NVIDIANemotron 3 UltraNVIDIA23NVIDIA
88阿里巴巴Qwen2-57B-A14B阿里巴巴23阿里巴巴
89Ling-3.0-flash-FinInclusionAI23InclusionAI
90Google DeepMindGemini 3.5 Flash-LiteGoogle DeepMind22Google DeepMind
91G9v3-39A5BAI9Stars22AI9Stars
92KAT-Coder-Pro V2KwaiKAT22KwaiKAT
93AlibabaQwen3.5 397B A17B (non-reasoning)Alibaba21Alibaba
94OpenAIGPT-6 Luna (low)OpenAI21OpenAI
95OpenAIGPT-5.6 Terra (non-reasoning)OpenAI21OpenAI
96K2 Horizon 7BInstitute of Foundation Models21Institute of Foundation Models
97阿里巴巴Qwen3.5-Omni-Plus阿里巴巴20阿里巴巴
98DeepSeekDeepSeek V4 Pro 0813 (non-reasoning)DeepSeek20DeepSeek
99OpenAIOpenAI o3OpenAI20OpenAI
100AlibabaQwen3.8 27B (non-reasoning)Alibaba20Alibaba
101Ling 3.0 FlashInclusionAI20InclusionAI
102K-EXAONE 2.0LG AI Research20LG AI Research
103LongCat 2.0LongCat19LongCat
104JT-35B-FlashChina Mobile19China Mobile
105阿里巴巴Qwen3.5-397B-A17B阿里巴巴18阿里巴巴
106OpenAIGPT-6 Luna (non-reasoning)OpenAI18OpenAI
107阿里巴巴Qwen3.6-35B-A3B阿里巴巴18阿里巴巴
108AlibabaQwen3.5 122B A10B (non-reasoning)Alibaba18Alibaba
109Facebook AI研究实验室Muse Glimmer-30B (high)Facebook AI研究实验室17Facebook AI研究实验室
110ByteDance SeedDoubao Seed CodeByteDance Seed17ByteDance Seed
111AnthropicHaiku 4.5Anthropic17Anthropic
112Google DeepMindGemma 4 26B A4BGoogle DeepMind17Google DeepMind
113Ring-2.6-1TInclusionAI17InclusionAI
114K2 Horizon 3.7BInstitute of Foundation Models16Institute of Foundation Models
115阿里巴巴Qwen3.5-122B-A10B阿里巴巴16阿里巴巴
116AnthropicHaiku 4.5 (non-reasoning)Anthropic15Anthropic
117AlibabaQwen3.6 35B A3B (non-reasoning)Alibaba15Alibaba
118Granite 4.2 30BIBM15IBM
119Google DeepMindGemma 4 31BGoogle DeepMind15Google DeepMind
120百度ERNIE 5.0 Thinking Preview百度14百度
121MistralAIMistral Medium 3.5MistralAI14MistralAI
122GoogleGemma 4 12BGoogle14Google
123亚马逊Nova 2 Pro(Preview) (medium)亚马逊14亚马逊
124GoogleGemma 4 31B (non-reasoning)Google14Google
125亚马逊Nova 2 Omni(Preview) (medium)亚马逊14亚马逊
126Apriel-v1.6-15B-ThinkerServiceNow13ServiceNow
127亚马逊Nova 2 Lite (high)亚马逊13亚马逊
128AlibabaQwen3.5 9B (non-reasoning)Alibaba13Alibaba
129EXAONE 4.5 33BLG AI Research13LG AI Research
130CohereCommand A+Cohere13Cohere
131GoogleGemma 4 26B A4B (non-reasoning)Google13Google
132AlibabaQwen3.5 4BAlibaba13Alibaba
133NVIDIANemotron 3.5 LightningNVIDIA13NVIDIA
134NVIDIANemotron 3 SuperNVIDIA13NVIDIA
135亚马逊Nova 2 Pro(Preview) (low)亚马逊13亚马逊
136亚马逊Nova 2 Lite (medium)亚马逊12亚马逊
137阿里巴巴Qwen3.5-Omni-Flash阿里巴巴12阿里巴巴
138MiniCPM5-2BOpenBMB12OpenBMB
139Mercury 2.5Inception12Inception
140JT-MINIChina Mobile12China Mobile
141MistralMagistral Medium 1.2Mistral12Mistral
142亚马逊Nova 2 Lite (low)亚马逊12亚马逊
143HyperNova 60B 2605 (high)Multiverse Computing12Multiverse Computing
144NVIDIANemotron Cascade 2 30B A3BNVIDIA12NVIDIA
145OpenAIGPT OSS 120B (high)OpenAI12OpenAI
146K2 Think V2MBZUAI11MBZUAI
147LongCat Flash LiteLongCat11LongCat
148HyperCLOVA X SEED Think (32B)Naver11Naver
149MistralMistral Small 4Mistral11Mistral
150阿里巴巴Qwen3-Next阿里巴巴11阿里巴巴
151阿里巴巴Qwen3.5-9B阿里巴巴11阿里巴巴
152亚马逊Nova 2 Omni(Preview) (low)亚马逊11亚马逊
153Granite 4.2 8BIBM11IBM
154Ling 3.0 TinyInclusionAI11InclusionAI
155Mi:dm K 2.5 ProKorea Telecom11Korea Telecom
156G9v3-3BAI9Stars11AI9Stars
157Trinity Large ThinkingArcee AI11Arcee AI
158AlibabaQwen3.5 4B (non-reasoning)Alibaba11Alibaba
159INTELLECT-3Prime Intellect11Prime Intellect
160Solar Open 100BUpstage10Upstage
161阿里巴巴Qwen3-Omni-30B-A3B阿里巴巴10阿里巴巴
162OpenAIGPT OSS 120B (low)OpenAI10OpenAI
163Facebook AI研究实验室Llama 4 MaverickFacebook AI研究实验室10Facebook AI研究实验室
164AmazonNova 2.0 Pro Preview (non-reasoning)Amazon10Amazon
165OpenAIGPT OSS 20B (low)OpenAI10OpenAI
166CohereNorth Mini CodeCohere10Cohere
167K2-V2 (high)MBZUAI10MBZUAI
168阿里巴巴Qwen3-Next阿里巴巴10阿里巴巴
169GoogleDiffusionGemma 26B A4BGoogle10Google
170GoogleGemma 4 12B (non-reasoning)Google9Google
171MistralAIMistral Large 3MistralAI9MistralAI
172阿里巴巴Qwen3-Coder-Next阿里巴巴9阿里巴巴
173Motif-2-12.7BMotif Technologies9Motif Technologies
174AmazonNova PremierAmazon9Amazon
175Granite 4.2 3BIBM9IBM
176K2-V2 (medium)MBZUAI9MBZUAI
177MistralMistral Small 4 (non-reasoning)Mistral9Mistral
178Tri-21B-ThinkTrillion Labs9Trillion Labs
179OpenAIGPT OSS 20B (high)OpenAI9OpenAI
180Google DeepMindGemma 4 E4BGoogle DeepMind9Google DeepMind
181NVIDIANemotron 3 NanoNVIDIA9NVIDIA
182OpenBMBMiniCPM5-1BOpenBMB9OpenBMB
183Sarvam 105B (high)Sarvam9Sarvam
184AmazonNova 2.0 Lite (non-reasoning)Amazon9Amazon
185MiniCPM5-1B (non-reasoning)OpenBMB9OpenBMB
186MistralMagistral Small 1.2Mistral9Mistral
187Nanbeige4.1-3BNanbeige8Nanbeige
188LFM2.5-2.6BLiquid AI8Liquid AI
189EXAONE 4.0 32BLG AI Research8LG AI Research
190AmazonNova 2.0 Omni (non-reasoning)Amazon8Amazon
191Facebook AI研究实验室Llama 4 ScoutFacebook AI研究实验室8Facebook AI研究实验室
192Hermes 4 70BNous Research8Nous Research
193Falcon-H1R-7BTII UAE8TII UAE
194Google DeepMindGemma 4 E2BGoogle DeepMind8Google DeepMind
195阿里巴巴Qwen3-Omni-30B-A3B阿里巴巴8阿里巴巴
196StepFunStep3 VL 10BStepFun8StepFun
197Facebook AI研究实验室Llama3.3-70B-InstructFacebook AI研究实验室8Facebook AI研究实验室
198NVIDIALlama Nemotron UltraNVIDIA8NVIDIA
199百度ERNIE-4.5-300B-A47B百度8百度
200Hermes 4 405BNous Research7Nous Research
201NVIDIANVIDIA Nemotron Nano 12B v2 VLNVIDIA7NVIDIA
202GoogleGemma 4 E4B (non-reasoning)Google7Google
203NVIDIANVIDIA Nemotron Nano 9B V2NVIDIA7NVIDIA
204Hermes 4 405B (non-reasoning)Nous Research7Nous Research
205NVIDIANemotron 3 Nano 4BNVIDIA7NVIDIA
206K2-V2 (low)MBZUAI7MBZUAI
207KimiKimi Linear 48B A3B InstructKimi7Kimi
208Facebook AI研究实验室Llama3.1-405BFacebook AI研究实验室7Facebook AI研究实验室
209LFM2.5-8B-A1BLiquid AI7Liquid AI
210Ring-flash-2.0InclusionAI7InclusionAI
211Olmo 3.1 32B ThinkAI27AI2
212CohereAIC4AI Command A (202503)CohereAI7CohereAI
213AlibabaQwen3.5 2BAlibaba7Alibaba
214NVIDIANemotron 3 Nano (non-reasoning)NVIDIA7NVIDIA
215NVIDIANVIDIA Nemotron Nano 9B V2 (non-reasoning)NVIDIA7NVIDIA
216Hermes 4 70B (non-reasoning)Nous Research7Nous Research
217Sarvam 30B (high)Sarvam7Sarvam
218Olmo 3.1 32B InstructAI26AI2
219GoogleGemma 4 E2B (non-reasoning)Google6Google
220PerplexityR1 1776Perplexity6Perplexity
221Facebook AI研究实验室Llama 3.2-Vision-90BFacebook AI研究实验室6Facebook AI研究实验室
222Celeris-1Celeris6Celeris
223Microsoft AzurePhi-4-mini-instruct (3.8B)Microsoft Azure6Microsoft Azure
224EXAONE 4.0 32B (non-reasoning)LG AI Research6LG AI Research
225AlibabaQwen3.5 2B (non-reasoning)Alibaba6Alibaba
226AlibabaQwen3.5 0.8BAlibaba6Alibaba
227DeepHermes 3 - Mistral 24B (non-reasoning)Nous Research6Nous Research
228Jamba 1.7 LargeAI21 Labs6AI21 Labs
229Granite 4.0 H SmallIBM6IBM
230MistralAIMinistral 3 14BMistralAI6MistralAI
231阿里巴巴Qwen3-Omni-30B-A3B阿里巴巴6阿里巴巴
232LFM2 24B A2BLiquid AI6Liquid AI
233Microsoft AzurePhi-4-reasoningMicrosoft Azure6Microsoft Azure
234亚马逊Amazon Nova Micro亚马逊6亚马逊
235NVIDIANVIDIA Nemotron Nano 12B v2 VL (non-reasoning)NVIDIA6NVIDIA
236Microsoft AzurePhi-4-multimodal-instruct Microsoft Azure6Microsoft Azure
237MiniCPM-V 4.6 1.3BOpenBMB6OpenBMB
238Jamba Reasoning 3BAI21 Labs6AI21 Labs
239Reka Flash 3Reka AI6Reka AI
240Olmo 3 7B ThinkAI26AI2
241Molmo 7B-DAllen Institute for AI6Allen Institute for AI
242MistralAIMinistral 3 8BMistralAI5MistralAI
243Facebook AI研究实验室Llama 3.2-Vision-11BFacebook AI研究实验室5Facebook AI研究实验室
244AlibabaQwen3.5 0.8B (non-reasoning)Alibaba5Alibaba
245Exaone 4.0 1.2BLG AI Research5LG AI Research
246Olmo 3 7BAI25AI2
247Exaone 4.0 1.2B (non-reasoning)LG AI Research5LG AI Research
248LFM2.5-1.2B-ThinkingLiquid AI5Liquid AI
249Jamba 1.7 MiniAI21 Labs5AI21 Labs
250LFM2.5-1.2B-InstructLiquid AI5Liquid AI
251Granite 4.0 H 1BIBM5IBM
252Google DeepMindGemma 3-270MGoogle DeepMind5Google DeepMind
253Apertus 70B InstructSwiss AI5Swiss AI
254Granite 4.0 MicroIBM5IBM
255DeepHermes 3 - Llama-3.1 8B (non-reasoning)Nous Research5Nous Research
256Molmo2-8BAI25AI2
257MistralMinistral 3 3BMistral5Mistral
258LFM2.5-VL-1.6BLiquid AI5Liquid AI
259Granite 4.0 350MIBM5IBM
260CohereTiny Aya GlobalCohere5Cohere
261Apertus 8B InstructSwiss AI5Swiss AI
262Granite 4.0 H 350MIBM5IBM
263K2 Horizon 0.9BInstitute of Foundation Models3Institute of Foundation Models

数据仅供参考,以官方来源为准。模型名称旁的链接可跳转到 DataLearner 模型详情页。

评测基准组成(Intelligence Index v4.0)

Intelligence Index 综合10项严格的评测基准,全面衡量AI模型能力,避免单一维度的过拟合。

GDPval-AA
智能体真实任务
τ²-Bench
智能体工具调用
Terminal-Bench
智能体编程
SciCode
编程能力
AA-LCR
长上下文推理
AA-Omniscience
知识与幻觉检测
IFBench
指令遵循
Humanity's Last Exam
推理与知识
GPQA Diamond
科学推理
CritPt
物理推理

常见问题 (FAQ)

什么是 Artificial Analysis Intelligence Index?▼
Artificial Analysis Intelligence Index v4.0 是一个综合评测指数,聚合了10项具有挑战性的评估——涵盖数学、科学、编程、智能体任务和推理——以全面衡量AI能力。它旨在防止单一维度的过拟合,提供一个统一分数来追踪模型进步。
智能指数是如何计算的?▼
该指数综合了10项评测的分数:GDPval-AA(智能体真实任务)、τ²-Bench(工具调用)、Terminal-Bench Hard(智能体编程)、SciCode(编程)、AA-LCR(长上下文推理)、AA-Omniscience(知识与幻觉检测)、IFBench(指令遵循)、Humanity's Last Exam(推理)、GPQA Diamond(科学推理)和 CritPt(物理推理)。所有测试由 Artificial Analysis 在标准化硬件上独立运行。
这与 LMArena 排行榜有什么区别?▼
LMArena 排名基于众包用户投票(盲测A/B对比的Elo评分),反映主观的人类偏好。而 Artificial Analysis Intelligence Index 使用标准化的自动评测基准进行客观评分,衡量特定领域的技术能力。两者各有价值——LMArena 捕捉真实用户体验,而 AA Intelligence Index 提供可复现的技术测量。
在哪里可以找到原始数据?▼
原始排行榜和详细方法论可在 artificialanalysis.ai 查看。Intelligence Index 的方法论详见 Intelligence Index 页面。