July 2026 · Research edition
AI model rankings Compare exact model configurations across independent user-preference, intelligence, objective-task, cost, and open-weight-model rankings—then test two on your own work.
Published Jul 24, 2026
Data cutoff Jul 24, 2026
Method OpenGPT research synthesis 1.1
Ranking basis 3 independent public sources Ranking basis
Full published ranking, verified official directory The main ranking preserves every publisher record, including entries whose exact official identity is not confirmed. The verified view and model directory keep only exact versions confirmed from provider documentation; missing data never becomes a zero score.
72 verified official versions 72
37 with public ranking data 37
35 awaiting comparable data 35
378 source records retained for audit 378 Model directory→
User preference score Intelligence index Objective-task score Cost per successful task Open-weight model score
Choose what to show Full source ranking Matched official versions
Show every publisher record, including aliases, run configurations, possible matches, and entries whose exact official version is still unconfirmed. Source ranks are preserved; ties can still cause rank numbers to skip.
User preference score LMArena’s complete latest style-controlled overall ranking. All 378 publisher records remain visible, with each record clearly marked as confirmed, possible, or not yet matched to an exact official version.
Independent source LMArena
Leaderboard publication date Jul 21, 2026
Ranking scope Text responses · Style-controlled · Overall ranking
Data retrieved Jul 24, 2026 Showing 40 of 378. User preference score.
Rank #1
Official version confirmed
Votes cast 14,646
License type Closed / proprietary Proprietary User preference score 1507 Based on anonymous head-to-head comparisons Likely score range 1,500.9–1,513.7
Rank #2
Official version confirmed
Model ID claude-opus-4-6
Votes cast 63,191
License type Closed / proprietary Proprietary User preference score 1505 Based on anonymous head-to-head comparisons Likely score range 1,501.1–1,508.6
Save+ Review match evidence↗ Model version confirmed, but site comparison cannot yet preserve this ranking configuration Rank #3
Official version confirmed
Model ID claude-opus-4-7
Votes cast 50,683
License type Closed / proprietary Proprietary User preference score 1502 Based on anonymous head-to-head comparisons Likely score range 1,497.8–1,506.2
Save+ Review match evidence↗ Model version confirmed, but site comparison cannot yet preserve this ranking configuration Rank #4
Official version confirmed
Votes cast 67,037
License type Closed / proprietary Proprietary User preference score 1498 Based on anonymous head-to-head comparisons Likely score range 1,494.1–1,501.4
Rank #5
muse-spark-1.1 Meta Exact model version not yet confirmed
Votes cast 7,927
License type Closed / proprietary Proprietary User preference score 1495 Based on anonymous head-to-head comparisons Likely score range 1,487.8–1,502.3
Rank #6
Official version confirmed
Votes cast 51,788
License type Closed / proprietary Proprietary User preference score 1494 Based on anonymous head-to-head comparisons Likely score range 1,489.5–1,497.9
Rank #7
muse-spark Meta Exact model version not yet confirmed
Votes cast 13,565
License type Closed / proprietary Proprietary User preference score 1488 Based on anonymous head-to-head comparisons Likely score range 1,481.7–1,493.4
Rank #8
Official version confirmed
Votes cast 84,631
License type Closed / proprietary Proprietary User preference score 1486 Based on anonymous head-to-head comparisons Likely score range 1,482.3–1,489.3
Rank #9
gemini-3-pro Google Exact model version not yet confirmed
Votes cast 41,268
License type Closed / proprietary Proprietary User preference score 1486 Based on anonymous head-to-head comparisons Likely score range 1,481.9–1,489.6
Rank #10
Official version confirmed
Votes cast 3,619
License type Closed / proprietary Proprietary User preference score 1486 Based on anonymous head-to-head comparisons Likely score range 1,475.7–1,495.6
Rank #11
Official version confirmed
Model ID gpt-5.6-sol
Votes cast 6,221
License type Closed / proprietary Proprietary User preference score 1485 Based on anonymous head-to-head comparisons Likely score range 1,477.2–1,493.0
Save+ Review match evidence↗ Model version confirmed, but site comparison cannot yet preserve this ranking configuration Rank #12
gemini-3.6-flash Google Exact model version not yet confirmed
Votes cast 4,747
License type Closed / proprietary Proprietary User preference score 1485 Based on anonymous head-to-head comparisons Likely score range 1,476.2–1,493.9
Rank #13
Official version confirmed
Model ID claude-opus-4-8
Votes cast 30,901
License type Closed / proprietary Proprietary User preference score 1484 Based on anonymous head-to-head comparisons Likely score range 1,478.5–1,488.7
Save+ Review match evidence↗ Model version confirmed, but site comparison cannot yet preserve this ranking configuration Rank #14
gpt-5.5-high OpenAI Exact model version not yet confirmed
Votes cast 45,760
License type Closed / proprietary Proprietary User preference score 1482 Based on anonymous head-to-head comparisons Likely score range 1,477.3–1,486.1
Rank #15
gpt-5.4-high OpenAI Exact model version not yet confirmed
Votes cast 58,997
License type Closed / proprietary Proprietary User preference score 1478 Based on anonymous head-to-head comparisons Likely score range 1,473.8–1,481.7
Rank #16
gpt-5.5 OpenAI Exact model version not yet confirmed
Votes cast 47,180
License type Closed / proprietary Proprietary User preference score 1476 Based on anonymous head-to-head comparisons Likely score range 1,472.0–1,480.7
Rank #17
Official version confirmed
Model ID gemini-3.5-flash
Votes cast 10,092
License type Closed / proprietary Proprietary User preference score 1476 Based on anonymous head-to-head comparisons Likely score range 1,469.7–1,482.8
Save+ Review match evidence↗ Model version confirmed, but site comparison cannot yet preserve this ranking configuration Rank #18
gpt-5.2-chat-latest-20260210 OpenAI Exact model version not yet confirmed
Votes cast 34,420
License type Closed / proprietary Proprietary User preference score 1476 Based on anonymous head-to-head comparisons Likely score range 1,471.6–1,479.8
Rank #19
qwen3.7-max-preview Alibaba Possible match: Qwen 3.7 Max · 2026-06-08 · exact version not confirmed
Votes cast 3,714
License type Closed / proprietary Proprietary User preference score 1475 Based on anonymous head-to-head comparisons Likely score range 1,465.1–1,485.2
Rank #20
grok-4.20-beta1 xAI Exact model version not yet confirmed
Votes cast 26,822
License type Closed / proprietary Proprietary User preference score 1474 Based on anonymous head-to-head comparisons Likely score range 1,469.7–1,479.0
Rank #21
Official version confirmed
Model ID gemini-3.5-flash
Votes cast 14,063
License type Closed / proprietary Proprietary User preference score 1474 Based on anonymous head-to-head comparisons Likely score range 1,468.2–1,480.4
Save+ Review match evidence↗ Model version confirmed, but site comparison cannot yet preserve this ranking configuration Rank #22
Official version confirmed
Votes cast 31,619
License type Closed / proprietary Proprietary User preference score 1473 Based on anonymous head-to-head comparisons Likely score range 1,468.3–1,478.5
Rank #23
gpt-5.5-instant OpenAI Exact model version not yet confirmed
Votes cast 25,995
License type Closed / proprietary Proprietary User preference score 1473 Based on anonymous head-to-head comparisons Likely score range 1,468.2–1,478.5
Rank #24
Official version confirmed
Model ID grok-4.20-0309-reasoning
Votes cast 60,330
License type Closed / proprietary Proprietary User preference score 1473 Based on anonymous head-to-head comparisons Likely score range 1,469.3–1,477.0
Rank #25
gemini-3-flash Google Exact model version not yet confirmed
Votes cast 30,682
License type Closed / proprietary Proprietary User preference score 1473 Based on anonymous head-to-head comparisons Likely score range 1,468.7–1,477.5
Rank #26
claude-opus-4-5-20251101-thinking-32k Anthropic Exact model version not yet confirmed
Votes cast 37,037
License type Closed / proprietary Proprietary User preference score 1473 Based on anonymous head-to-head comparisons Likely score range 1,469.0–1,476.8
Rank #27
claude-sonnet-4-6 Anthropic Exact model version not yet confirmed
Votes cast 57,183
License type Closed / proprietary Proprietary User preference score 1473 Based on anonymous head-to-head comparisons Likely score range 1,468.7–1,476.4
Rank #28
grok-4.20-multi-agent-beta-0309 xAI Exact model version not yet confirmed
Votes cast 59,116
License type Closed / proprietary Proprietary User preference score 1471 Based on anonymous head-to-head comparisons Likely score range 1,467.1–1,474.8
Rank #29
glm-5.1 Z.AI Exact model version not yet confirmed
Votes cast 30,726
License type MIT User preference score 1470 Based on anonymous head-to-head comparisons Likely score range 1,465.1–1,474.2
Rank #30
glm-5.2 (max) Z.AI Exact model version not yet confirmed
Votes cast 18,017
License type MIT User preference score 1469 Based on anonymous head-to-head comparisons Likely score range 1,463.4–1,475.0
Rank #31
claude-opus-4-5-20251101 Anthropic Exact model version not yet confirmed
Votes cast 70,976
License type Closed / proprietary Proprietary User preference score 1469 Based on anonymous head-to-head comparisons Likely score range 1,465.9–1,472.3
Rank #32
ernie-5.1 Baidu Exact model version not yet confirmed
Votes cast 37,418
License type Closed / proprietary Proprietary User preference score 1468 Based on anonymous head-to-head comparisons Likely score range 1,463.1–1,472.4
Rank #33
Official version confirmed
Votes cast 7,833
License type Closed / proprietary Proprietary User preference score 1468 Based on anonymous head-to-head comparisons Likely score range 1,460.4–1,475.0
Rank #34
mimo-v2.5-pro Xiaomi Exact model version not yet confirmed
Votes cast 42,195
License type MIT User preference score 1467 Based on anonymous head-to-head comparisons Likely score range 1,462.4–1,471.3
Rank #35
gpt-5.4 OpenAI Exact model version not yet confirmed
Votes cast 62,036
License type Closed / proprietary Proprietary User preference score 1466 Based on anonymous head-to-head comparisons Likely score range 1,462.4–1,470.1
Rank #36
grok-4.1-thinking xAI Exact model version not yet confirmed
Votes cast 65,461
License type Closed / proprietary Proprietary User preference score 1466 Based on anonymous head-to-head comparisons Likely score range 1,462.7–1,469.1
Rank #37
qwen3.5-max-preview Alibaba Exact model version not yet confirmed
Votes cast 21,479
License type Closed / proprietary Proprietary User preference score 1465 Based on anonymous head-to-head comparisons Likely score range 1,459.8–1,469.8
Rank #38
Official version confirmed
Model ID claude-sonnet-5
Votes cast 13,521
License type Closed / proprietary Proprietary User preference score 1461 Based on anonymous head-to-head comparisons Likely score range 1,454.9–1,467.4
Save+ Review match evidence↗ Model version confirmed, but site comparison cannot yet preserve this ranking configuration Rank #39
Official version confirmed
Votes cast 37,686
License type Modified MIT User preference score 1461 Based on anonymous head-to-head comparisons Likely score range 1,456.4–1,465.5
Rank #40
qwen3.6-max-preview Alibaba Exact model version not yet confirmed
Votes cast 5,188
License type Closed / proprietary Proprietary User preference score 1460 Based on anonymous head-to-head comparisons Likely score range 1,451.7–1,468.5
Showing 40 of 378 Show more Show all
Ranking basis
How to read this research edition 01 Scores stay on their original publisher scales. They are not combined into a composite score.
02 Missing data does not mean zero. It means the publisher did not report that metric for this snapshot.
03 When a source reports model settings, the position applies to that exact setup and source release—not every task, region, or deployment.
04 No provider can pay for placement. OpenGPT links directly to the evidence used.
Ranking basis
Data and sources Each view preserves the source’s metric, release, and methodology. Follow the links to audit the underlying work. AR LMArena The complete latest style-controlled overall ranking: all 378 source-listed model entries, based on users comparing two model answers side by side.
Leaderboard publication date Jul 21, 2026
Data retrieved Jul 24, 2026
License CC BY 4.0 Original ranking ↗ Scoring method ↗
AA Artificial Analysis Nine evaluations across agentic work, coding, general knowledge, and scientific reasoning.
Source snapshot Intelligence Index v4.1 · continuously updated
Data retrieved Jul 24, 2026 Original ranking ↗ Scoring method ↗
LB LiveBench Twenty-three objective tasks across seven categories; questions refresh every six months.
Source snapshot LiveBench-2026-06-25
Data retrieved Jul 24, 2026 Original ranking ↗ Scoring method ↗
Method
A predictable publishing rhythm OpenGPT will separate changing signals from frozen editions so past decisions remain auditable. 01 Weekly watch Review new model and benchmark releases; flag material changes without silently rewriting this edition.
02 Monthly frozen edition Publish a dated snapshot with exact configurations, source releases, and corrections.
03 Quarterly method review Review categories, source quality, conflicts, and inclusion rules before changing the method.
Current July 2026 · Research edition First published edition · Download frozen data
August 2026 August 2026 Scheduled next edition. It will be published only after source capture and review.