Skip to content
TrustList
AI Models

AI Model Rankings

An independent leaderboard — models scored on a hybrid of public benchmarks, verified reviews and community votes. Transparent by design: every model links to its full profile.

Vote#ModelScoreCapabilityRatingCursorBench 3.2.0DeepSWEFrontierCode 1.1 MainNL2RepoSWE-bench Verified
1
Anthropic · 1,000,000 ctx · Proprietary
96
+30 vs avg
9696 %
+29.95
295
391
+25 vs avg
8988.6 %
+22.55
4
Anthropic · 1,000,000 ctx · Proprietary
81
+15 vs avg
8282.1 %
+16.05
579
6
OpenAI · 200,000 ctx · Proprietary
75
+9 vs avg
7271.7 %
+5.65
7
Anthropic · 1,000,000 ctx · Proprietary
73
+7 vs avg
7373.4 %
8
Google · 1,000,000 ctx · Proprietary
67
+1 vs avg
6463.8 %
2.25
9
DeepSeek · 1,000,000 ctx · Proprietary
62
4 vs avg
6262.7 %
4.73
61.5 %
1061
1161
1260
7 vs avg
7474.3 %
+6.87
13
IBM · 30B params · 512,000 ctx · Apache 2.0
57
9 vs avg
5757 %
9.05
1457
9 vs avg
5554.6 %
11.45
1555
16
Google · 1,000,000 ctx · Proprietary
55
12 vs avg
5565.3 %
2.13
43.6 %
1751
18
IBM · 8B params · 512,000 ctx · Apache 2.0
48
18 vs avg
4847.67 %
18.38
1945
2042
2141
2239
23
OpenAI · 128,000 ctx · Proprietary
39
27 vs avg
3333 %
33.05
2436
2534
2633
2733
2832
2931
3031
3131
3229
3329
3428
3527
3625
3724
3824
3922
4021
4121
4219
4318
4417
4517
4615
4715
4815
4914
5014
5114
5213
5312
5412
5512
5611
5711
5810
5910
609
619
628
637
647
657
667
676
686
695
70
Anthropic · 1,000,000 ctx · Proprietary
0
71
Anthropic · Proprietary
0
72
Anthropic · 1,000,000 ctx · Proprietary
0
73
DeepSeek · 1,000,000 ctx · Proprietary
0
74
DeepSeek · 1,000,000 ctx
0
750
760
770
78
Google · 1,000,000 ctx · Proprietary
0
79
Google · Proprietary
0
80
Google · Restricted — Fairwind Program
0
81
Z.ai (Zhipu AI) · 1,000,000 ctx · Proprietary at launch; weights released ~2 weeks later
0
82
Z.ai (Zhipu AI) · 1,000,000 ctx · Open weights
0
83
0
840
85
OpenAI · 1,050,000 ctx · Proprietary
0
86
IBM · 30B params · 512,000 ctx · Apache 2.0
0
87
IBM · 8B params · 512,000 ctx · Apache 2.0
0
88
xAI · 500,000 ctx · Proprietary
0
89
Tencent · 770B total, 49B active (mixture of experts) params · 1,000,000 ctx
0
900
91
Moonshot AI · 2.8T (104B active) params · 1,048,576 ctx · Kimi K3 License (open weights, not OSI)
0
92
Meta · 1,000,000 ctx · Proprietary
0
93
Alibaba · 1,000,000 ctx · Proprietary
0
94
Alibaba · 27B params · 1,000,000 ctx · Qwen (open weights)
0
95
Alibaba · 2.4T (95B active) params · 1,000,000 ctx · Proprietary (weights announced)
0
96
Alibaba Cloud (Qwen) · 1,000,000 ctx
0
97
Google DeepMind · Proprietary (Google Cloud access)
0

Score = hybrid of benchmark capability (60%), review rating (25%) and community votes (15%), renormalized by available signals. Capability is the mean of normalized benchmark results; green/red deltas compare each value to the average across benchmarked models. Official (editor-verified) and community-submitted values are kept separate — community submissions only affect rankings after editor approval. Your vote, rating and submitted stats all feed this score — sign-in required, so rankings stay hard to game.

Missing a model?

Suggest one and our editors will review and add it to the board.

Sign in to suggest a model
Launches

New AI model launches

The models that launched on TrustList this month, voted up by the community.

All AI launches