AI Model Benchmarks

733 scores across 9 benchmarks — quality, coding, reasoning, and speed.

Every score we hold, including models older than 12 months. For current picks see the rankings, which use the last 12 months only.

Output throughput in tokens per second measured on standard prompts. Higher = faster responses. Source: artificialanalysis.ai and vendor-published measurements.

Model
tok/s
1
1012
15
215
18
201
21
192
24
184
25
176
26
175
28
163
31
158
33
155
36
145
38
138
40
136
44
119
45
119
46
118
47
115
48
115
49
114
50
114
51
113
52
112
54
108
58
98
61
92
68
80
69
78
77
70
81
69
85
66
88
63
89
62
95
60
102
55
104
51
111
42
112
42
118
37
120
32
121
28
125
22

Quality scores from official model cards, published technical reports, and independent leaderboards (vals.ai, Vectara). Speed benchmarks from artificialanalysis.ai (approximate medians). Higher is better for all metrics except Hallucination Rate, where lower is better.