Skip to content
TrustList
AI Models

AI Model Rankings

An independent leaderboard — models scored on a hybrid of public benchmarks, verified reviews and community votes. Transparent by design: every model links to its full profile.

Vote#ModelScoreCapabilityRatingARC-AGI-3 (Provider Adapter harness)ARC-AGI-3 (Standard harness)Artificial Analysis Intelligence IndexAutomationBenchChartographyCursorBench 3.2.0
1
Anthropic · 1,000,000 ctx · Proprietary
96
+25 vs avg
96
289
+17 vs avg
86
388
+17 vs avg
87
4
Anthropic · 1,000,000 ctx · Proprietary
87
+15 vs avg
89
5
OpenAI · 1,050,000 ctx · Proprietary
87
+15 vs avg
8799.9 %62.7 %
6
Alibaba · 2.4T (95B active) params · 1,000,000 ctx · Proprietary (weights announced)
86
+15 vs avg
86
7
Z.ai (Zhipu AI) · 1,000,000 ctx · Proprietary at launch; weights released ~2 weeks later
85
+13 vs avg
85
8
OpenAI · 200,000 ctx · Proprietary
81
+10 vs avg
80
979
10
Google · 1,000,000 ctx · Proprietary
78
+6 vs avg
77
11
Google · Proprietary
73
+1 vs avg
73
1272
+1 vs avg
75
1371
+0 vs avg
84
1471
+0 vs avg
72
15
DeepSeek · 1,000,000 ctx · Proprietary
71
1 vs avg
71
1669
3 vs avg
72
17
Google · 1,000,000 ctx · Proprietary
64
8 vs avg
64
18
OpenAI · 128,000 ctx · Proprietary
62
10 vs avg
61
19
xAI · 500,000 ctx · Proprietary
61
10 vs avg
6161 index
2061
21
IBM · 30B params · 512,000 ctx · Apache 2.0
57
14 vs avg
57
22
Anthropic · 1,000,000 ctx · Proprietary
53
18 vs avg
5331.4 %73.4 %
2351
2449
22 vs avg
6148.3 %
25
IBM · 8B params · 512,000 ctx · Apache 2.0
48
24 vs avg
48
26
Google · Restricted — Fairwind Program
47
24 vs avg
47
2745
2842
2941
3039
3136
3234
3333
3433
3532
3631
3731
3831
3929
4029
4128
4227
4325
4424
4524
4622
4721
4819
4918
5017
5117
5215
5315
5415
5514
5614
5714
5813
5912
6012
6112
6211
6311
6410
6510
669
679
688
697
707
717
727
736
746
755
760
77
Anthropic · 1,000,000 ctx · Proprietary
0
78
Anthropic · Proprietary
0
79
Anthropic · Proprietary
0
80
Anthropic · 1,000,000 ctx · Proprietary
0
81
DeepSeek · 1,000,000 ctx · Proprietary
0
82
DeepSeek · 1,000,000 ctx
0
83
DeepSeek · 1,000,000 ctx
0
840
850
860
87
Google · 1,000,000 ctx · Proprietary
0
88
Google · Proprietary
0
89
Google · Restricted — Fairwind Program
0
90
Z.ai (Zhipu AI) · 1,000,000 ctx · Proprietary at launch; weights released ~2 weeks later
0
91
Z.ai (Zhipu AI) · 1,000,000 ctx · Open weights
0
92
Z.ai (Zhipu AI) · 1,000,000 ctx · Open weights
0
93
0
94
0
950
960
97
OpenAI · 1,050,000 ctx · Proprietary
0
98
IBM · 30B params · 512,000 ctx · Apache 2.0
0
99
IBM · 8B params · 512,000 ctx · Apache 2.0
0
100
xAI · 500,000 ctx · Proprietary
0
101
Tencent · 770B total, 49B active (mixture of experts) params · 1,000,000 ctx
0
102
Tencent · 770B total, 49B active (mixture of experts) params · 1,000,000 ctx
0
1030
104
Moonshot AI · 2.8T (104B active) params · 1,048,576 ctx · Kimi K3 License (open weights, not OSI)
0
105
Moonshot AI · 2.8T (104B active) params · 1,048,576 ctx · Kimi K3 License (open weights, not OSI)
0
106
Meta · 1,000,000 ctx · Proprietary
0
107
Meta · 1,000,000 ctx · Proprietary
0
108
Alibaba · 1,000,000 ctx · Proprietary
0
109
Alibaba · 1,000,000 ctx · Proprietary
0
110
Alibaba · 27B params · 1,000,000 ctx · Qwen (open weights)
0
111
Alibaba · 27B params · 1,000,000 ctx · Qwen (open weights)
0
112
Alibaba · 2.4T (95B active) params · 1,000,000 ctx · Proprietary (weights announced)
0
113
Alibaba Cloud (Qwen) · 1,000,000 ctx
0
114
Alibaba Cloud (Qwen) · 1,000,000 ctx
0
115
Google DeepMind · Proprietary (Google Cloud access)
0
116
Google DeepMind · Proprietary (Google Cloud access)
0

Score = hybrid of benchmark capability (60%), review rating (25%) and community votes (15%), renormalized by available signals. Capability is the mean of normalized benchmark results; green/red deltas compare each value to the average across benchmarked models. Official (editor-verified) and community-submitted values are kept separate — community submissions only affect rankings after editor approval. Your vote, rating and submitted stats all feed this score — sign-in required, so rankings stay hard to game.

Missing a model?

Suggest one and our editors will review and add it to the board.

Sign in to suggest a model
Launches

New AI model launches

The models that launched on TrustList this month, voted up by the community.

All AI launches