Rankings

Model rankings

Open models ranked by intelligence and speed, hosted on GPUs in the EU.

Bench: measured quality of our models

The 50 best models we host, measured numbers only (Epoch AI, or LiveBench and LMArena translated onto the same scale); estimates are not ranked.

# Model Source Range LiveBench Bench
1
Moonshotai
Measured 81.4-88.3 79.2
84.6
2
Z.AI
Measured 79.5-85.0 76.1
82.0
3 Measured 79.6-84.1 77.4
81.7
4 Measured 73.2-84.4 81.1
81.3
5 Via LiveBench 71.3-83.9 76.8
77.6
6 Measured 74.5-80.3 71.6
77.4
7
Z.AI
Measured 74.9-79.5 73.2
77.2
8 Via LiveBench 70.9-83.5 76.2
77.2
9
Moonshotai
Measured 73.9-78.5 70.5
76.3
10
Moonshotai
Measured 72.7-77.1 68.4
75.0
11
Z.AI
Measured 72.4-77.0 70.2
74.8
12 Measured 71.8-76.6 75.3
74.2
13
DeepSeek
Measured 71.8-76.0 71.6
73.9
14
Moonshotai
Measured 70.3-74.1 69.1
72.6
15
Google
Via LMArena 63.8-79.7 —
71.8
16
MiniMaxAI
Measured 66.2-74.6 67.3
71.3
17
MiniMaxAI
Measured 65.9-72.3 60.2
70.9
18
Qwen3.5-397b-a17b
Measured 68.5-72.8 —
70.8
19 Measured 67.6-72.2 64.0
70.6
20
DeepSeek
Measured 67.5-71.8 51.8
70.4
21 Measured 66.8-72.2 74.2
70.1
22
Moonshotai
Measured 66.2-71.8 61.6
70.0
23 Via LiveBench 63.6-76.2 67.4
69.9
24
Z.AI
Measured 67.2-72.1 68.9
69.8
25 Measured 65.4-71.2 49.9
68.8
26 Via LMArena 60.3-76.2 —
68.3
27
Z.AI
Via LiveBench 61.6-74.3 65.0
68.0
28 Via LMArena 59.9-75.8 —
67.8
29 Via LMArena 59.9-75.8 —
67.8
30 Via LiveBench 61.4-74.0 64.7
67.7
31 Measured 64.0-70.1 —
67.4
32
Z.AI
Measured 63.8-69.4 58.1
66.9
33
Qwen
Via LMArena 58.6-74.5 —
66.6
34 Via LMArena 58.6-74.5 —
66.6
35 Via LMArena 58.6-74.5 —
66.6
36
Google
Measured 62.2-68.6 61.6
65.9
37 Measured 62.6-68.4 —
65.7
38 Measured 60.3-66.9 —
64.8
39 Via LiveBench 58.0-70.6 60.5
64.3
40
DeepSeek
Measured 61.0-66.4 69.4
64.1
41
Moonshotai
Measured 57.6-65.3 48.1
62.7
42
Gpt-oss-120b
Measured 56.7-65.6 46.1
62.4
43
DeepSeek
Measured 56.6-66.5 —
62.4
44
Qwen
Measured 57.8-64.1 —
61.8
45 Measured 56.6-63.5 63.4
61.7
46
DeepSeek
Measured 57.5-63.0 72.7
61.2
47
Qwen3-235b-a22b-instruct-2507
Measured 56.8-63.5 48.8
61.2
48 Via LiveBench 54.8-67.4 56.6
61.1
49
Qwen
Measured 55.9-63.1 43.6
60.6
50
Qwen
Measured 54.2-62.6 68.2
60.3

Sources: Epoch AI and LMArena (CC BY 4.0), LiveBench. Refreshed every morning. epoch.ai ↗

Intelligence leaderboard

Ranked by a composite of publicly reported benchmarks. Higher is better.

# Model Size MMLUGPQAMATHCode Index
1
DeepSeek
— 91.4 81.0 97.0 92.0
90.4
2
Gpt-oss-120b
— — 80.1 97.9 —
89.0
3
DeepSeek
— 90.8 71.5 97.3 90.0
87.4
4
Z.AI
— 85.5 82.0 93.5 —
87.0
5
Z.AI
— 84.6 79.1 91.0 —
84.9
6 32B 83.3 — 83.1 88.4
84.9
7
OpenAI
21B — 71.5 96.0 —
83.8
8 235B 87.8 71.1 92.2 —
83.7
9
DeepSeek
— 90.0 67.0 92.0 85.0
83.5
10 685B 89.5 62.7 91.6 84.0
82.0
11 110B 81.0 75.0 89.0 —
81.7
12 14B 79.7 — 80.0 83.5
81.1
13
DeepSeek
— 88.5 59.1 90.2 82.6
80.1
14 70B — 65.2 94.5 —
79.9
15
Qwen
32B 82.4 68.4 88.5 —
79.8
16 32B — 62.1 94.3 —
78.2
17
Qwen
14B 80.7 65.8 87.7 —
78.1
18 30B 79.5 65.8 86.6 —
77.3
19 14B — 59.1 93.9 —
76.5
20 72B 86.1 49.0 83.1 86.6
76.2
21
Microsoft
14B 84.8 56.1 80.4 82.6
76.0
22 70B 86.0 50.5 77.0 88.4
75.5
23 405B 87.3 50.7 73.8 89.0
75.2
24 32B — — 57.2 92.7
75.0
25
Qwen
8B 76.9 62.0 85.0 —
74.6
26 7B — — 57.2 88.4
72.8
27 7.6B — 49.1 92.8 —
71.0
28 70B 83.6 41.7 68.0 80.5
68.5
29
Qwen
7B 74.2 36.4 75.5 84.8
67.7
30
Microsoft
3.8B 67.3 — 64.0 —
65.7
31
Google
27B — 42.4 89.0 —
65.7
32
Google
12B — 34.9 83.8 —
59.4
33 27B 75.2 — 42.3 —
58.8
34 3.8B 69.0 — 48.5 —
58.8
35 8B 69.4 30.4 51.9 72.6
56.1
36 3B 63.4 — 48.0 —
55.7
37
Google
9.2B 71.3 — 36.6 —
54.0
38
Google
4B — 30.8 75.6 —
53.2
39
Mistral
12B 68.0 — 38.0 —
53.0
40 1B 49.3 — 30.6 —
40.0
41 7.2B 60.1 — 11.2 30.5
33.9

Benchmarks are publicly reported figures from model authors and open leaderboards (last updated 2026-06). Methodology varies per model, so the index is an ordering aid, not a controlled benchmark.

Speed leaderboard

Ranked by the real time-to-respond we measured on our own EU GPUs. Lower is better.

# Model Size Response
1
Google
2.5B 66 ms
2
Google
2.5B 67 ms
3 3.2B 76 ms
4 1.5B 81 ms
5 6.2B 81 ms
6 3.8B 81 ms
7 3.8B 82 ms
8 3B 83 ms
9 1.2B 87 ms
10 8B 89 ms
11 4.1B 93 ms
12 2.1B 93 ms
13 6.2B 94 ms
14
Qwen
2B 96 ms
15 0.9B 96 ms
16 15B 98 ms
17
Microsoft
15B 98 ms
18 8.3B 101 ms
19 7.8B 102 ms
20 7.8B 103 ms
21 3.8B 103 ms
22
Microsoft
8.3B 106 ms
23 23B 109 ms
24
LumiOpen
8B 109 ms
25
Utter-project
2.1B 110 ms

Response times were measured at our last verification of each model and vary with load and cold starts.

Host. Route. Ship.

No credit card required. Pay as you go, cancel anytime.

Start Hosting Free Today