AI Model Benchmarks

854 scores across 9 benchmarks — quality, coding, reasoning, and speed.

Every score we hold, including models older than 12 months. For current picks see the rankings, which use the last 12 months only.

Output throughput in tokens per second measured on standard prompts. Higher = faster responses. Source: artificialanalysis.ai and vendor-published measurements.

Model
tok/s
3
699
22
215
27
201
30
192
34
184
37
175
40
163
44
158
45
155
50
145
52
138
55
136
56
134
58
129
59
127
64
115
65
114
66
114
68
113
69
112
70
108
72
106
75
106
77
104
79
98
86
92
92
87
101
81
102
81
105
78
108
74
109
73
113
70
115
70
116
69
118
66
126
62
128
60
129
60
131
60
132
59
135
58
136
57
137
57
140
55
142
53
146
50
148
50
150
49
155
45
156
43.167
157
43
158
42
159
41
168
33
169
32
170
28.103
171
28
172
28
174
23.904
175
23
176
22

Quality scores from official model cards, published technical reports, and independent leaderboards (vals.ai, Vectara). Speed benchmarks from artificialanalysis.ai (approximate medians). Higher is better for all metrics except Hallucination Rate, where lower is better.