July 2026 · Research edition

AI model rankings

Compare exact model configurations across independent user-preference, intelligence, objective-task, cost, and open-weight-model rankings—then test two on your own work.

Published
Jul 24, 2026
Data cutoff
Jul 24, 2026
Method
OpenGPT research synthesis 1.1
Ranking basis
3 independent public sources

Ranking basis

Full published ranking, verified official directory

The main ranking preserves every publisher record, including entries whose exact official identity is not confirmed. The verified view and model directory keep only exact versions confirmed from provider documentation; missing data never becomes a zero score.
72 verified official versions
72
37 with public ranking data
37
35 awaiting comparable data
35
378 source records retained for audit
378
Model directory

July 2026 · Research edition

Choose a ranking metric

378 source-listed model entries

Choose what to show

Show every publisher record, including aliases, run configurations, possible matches, and entries whose exact official version is still unconfirmed. Source ranks are preserved; ties can still cause rank numbers to skip.

User preference score

LMArena’s complete latest style-controlled overall ranking. All 378 publisher records remain visible, with each record clearly marked as confirmed, possible, or not yet matched to an exact official version.

Independent sourceLMArena

Showing 40 of 378. User preference score.

  1. Rank #1

    Official version confirmed

    Votes cast
    14,646
    License type
    Closed / proprietaryProprietary
    User preference score1507Based on anonymous head-to-head comparisons
    Likely score range1,500.91,513.7
    Review match evidence
  2. Rank #2

    Official version confirmed

    Model ID
    claude-opus-4-6
    Votes cast
    63,191
    License type
    Closed / proprietaryProprietary
    User preference score1505Based on anonymous head-to-head comparisons
    Likely score range1,501.11,508.6
    Review match evidenceModel version confirmed, but site comparison cannot yet preserve this ranking configuration
  3. Rank #3

    Official version confirmed

    Model ID
    claude-opus-4-7
    Votes cast
    50,683
    License type
    Closed / proprietaryProprietary
    User preference score1502Based on anonymous head-to-head comparisons
    Likely score range1,497.81,506.2
    Review match evidenceModel version confirmed, but site comparison cannot yet preserve this ranking configuration
  4. Rank #4

    Official version confirmed

    Votes cast
    67,037
    License type
    Closed / proprietaryProprietary
    User preference score1498Based on anonymous head-to-head comparisons
    Likely score range1,494.11,501.4
    Review match evidence
  5. Rank #5

    muse-spark-1.1

    Meta

    Exact model version not yet confirmed

    Votes cast
    7,927
    License type
    Closed / proprietaryProprietary
    User preference score1495Based on anonymous head-to-head comparisons
    Likely score range1,487.81,502.3
  6. Rank #6

    Official version confirmed

    Votes cast
    51,788
    License type
    Closed / proprietaryProprietary
    User preference score1494Based on anonymous head-to-head comparisons
    Likely score range1,489.51,497.9
    Review match evidence
  7. Rank #7

    muse-spark

    Meta

    Exact model version not yet confirmed

    Votes cast
    13,565
    License type
    Closed / proprietaryProprietary
    User preference score1488Based on anonymous head-to-head comparisons
    Likely score range1,481.71,493.4
  8. Rank #8

    Official version confirmed

    Votes cast
    84,631
    License type
    Closed / proprietaryProprietary
    User preference score1486Based on anonymous head-to-head comparisons
    Likely score range1,482.31,489.3
    Review match evidence
  9. Rank #9

    gemini-3-pro

    Google

    Exact model version not yet confirmed

    Votes cast
    41,268
    License type
    Closed / proprietaryProprietary
    User preference score1486Based on anonymous head-to-head comparisons
    Likely score range1,481.91,489.6
  10. Rank #10

    kimi-k3

    Moonshot AI

    Official version confirmed

    Votes cast
    3,619
    License type
    Closed / proprietaryProprietary
    User preference score1486Based on anonymous head-to-head comparisons
    Likely score range1,475.71,495.6
    Review match evidence
  11. Rank #11

    Official version confirmed

    Model ID
    gpt-5.6-sol
    Votes cast
    6,221
    License type
    Closed / proprietaryProprietary
    User preference score1485Based on anonymous head-to-head comparisons
    Likely score range1,477.21,493.0
    Review match evidenceModel version confirmed, but site comparison cannot yet preserve this ranking configuration
  12. Rank #12

    gemini-3.6-flash

    Google

    Exact model version not yet confirmed

    Votes cast
    4,747
    License type
    Closed / proprietaryProprietary
    User preference score1485Based on anonymous head-to-head comparisons
    Likely score range1,476.21,493.9
  13. Rank #13

    Official version confirmed

    Model ID
    claude-opus-4-8
    Votes cast
    30,901
    License type
    Closed / proprietaryProprietary
    User preference score1484Based on anonymous head-to-head comparisons
    Likely score range1,478.51,488.7
    Review match evidenceModel version confirmed, but site comparison cannot yet preserve this ranking configuration
  14. Rank #14

    gpt-5.5-high

    OpenAI

    Exact model version not yet confirmed

    Votes cast
    45,760
    License type
    Closed / proprietaryProprietary
    User preference score1482Based on anonymous head-to-head comparisons
    Likely score range1,477.31,486.1
  15. Rank #15

    gpt-5.4-high

    OpenAI

    Exact model version not yet confirmed

    Votes cast
    58,997
    License type
    Closed / proprietaryProprietary
    User preference score1478Based on anonymous head-to-head comparisons
    Likely score range1,473.81,481.7
  16. Rank #16

    gpt-5.5

    OpenAI

    Exact model version not yet confirmed

    Votes cast
    47,180
    License type
    Closed / proprietaryProprietary
    User preference score1476Based on anonymous head-to-head comparisons
    Likely score range1,472.01,480.7
  17. Rank #17

    Official version confirmed

    Model ID
    gemini-3.5-flash
    Votes cast
    10,092
    License type
    Closed / proprietaryProprietary
    User preference score1476Based on anonymous head-to-head comparisons
    Likely score range1,469.71,482.8
    Review match evidenceModel version confirmed, but site comparison cannot yet preserve this ranking configuration
  18. Rank #18

    gpt-5.2-chat-latest-20260210

    OpenAI

    Exact model version not yet confirmed

    Votes cast
    34,420
    License type
    Closed / proprietaryProprietary
    User preference score1476Based on anonymous head-to-head comparisons
    Likely score range1,471.61,479.8
  19. Rank #19

    qwen3.7-max-preview

    Alibaba

    Possible match: Qwen 3.7 Max · 2026-06-08 · exact version not confirmed

    Votes cast
    3,714
    License type
    Closed / proprietaryProprietary
    User preference score1475Based on anonymous head-to-head comparisons
    Likely score range1,465.11,485.2
  20. Rank #20

    grok-4.20-beta1

    xAI

    Exact model version not yet confirmed

    Votes cast
    26,822
    License type
    Closed / proprietaryProprietary
    User preference score1474Based on anonymous head-to-head comparisons
    Likely score range1,469.71,479.0
  21. Rank #21

    Official version confirmed

    Model ID
    gemini-3.5-flash
    Votes cast
    14,063
    License type
    Closed / proprietaryProprietary
    User preference score1474Based on anonymous head-to-head comparisons
    Likely score range1,468.21,480.4
    Review match evidenceModel version confirmed, but site comparison cannot yet preserve this ranking configuration
  22. Rank #22

    Official version confirmed

    Votes cast
    31,619
    License type
    Closed / proprietaryProprietary
    User preference score1473Based on anonymous head-to-head comparisons
    Likely score range1,468.31,478.5
    Review match evidence
  23. Rank #23

    gpt-5.5-instant

    OpenAI

    Exact model version not yet confirmed

    Votes cast
    25,995
    License type
    Closed / proprietaryProprietary
    User preference score1473Based on anonymous head-to-head comparisons
    Likely score range1,468.21,478.5
  24. Rank #24

    Official version confirmed

    Model ID
    grok-4.20-0309-reasoning
    Votes cast
    60,330
    License type
    Closed / proprietaryProprietary
    User preference score1473Based on anonymous head-to-head comparisons
    Likely score range1,469.31,477.0
    Review match evidence
  25. Rank #25

    gemini-3-flash

    Google

    Exact model version not yet confirmed

    Votes cast
    30,682
    License type
    Closed / proprietaryProprietary
    User preference score1473Based on anonymous head-to-head comparisons
    Likely score range1,468.71,477.5
  26. Rank #26

    claude-opus-4-5-20251101-thinking-32k

    Anthropic

    Exact model version not yet confirmed

    Votes cast
    37,037
    License type
    Closed / proprietaryProprietary
    User preference score1473Based on anonymous head-to-head comparisons
    Likely score range1,469.01,476.8
  27. Rank #27

    claude-sonnet-4-6

    Anthropic

    Exact model version not yet confirmed

    Votes cast
    57,183
    License type
    Closed / proprietaryProprietary
    User preference score1473Based on anonymous head-to-head comparisons
    Likely score range1,468.71,476.4
  28. Rank #28

    grok-4.20-multi-agent-beta-0309

    xAI

    Exact model version not yet confirmed

    Votes cast
    59,116
    License type
    Closed / proprietaryProprietary
    User preference score1471Based on anonymous head-to-head comparisons
    Likely score range1,467.11,474.8
  29. Rank #29

    glm-5.1

    Z.AI

    Exact model version not yet confirmed

    Votes cast
    30,726
    License type
    MIT
    User preference score1470Based on anonymous head-to-head comparisons
    Likely score range1,465.11,474.2
  30. Rank #30

    glm-5.2 (max)

    Z.AI

    Exact model version not yet confirmed

    Votes cast
    18,017
    License type
    MIT
    User preference score1469Based on anonymous head-to-head comparisons
    Likely score range1,463.41,475.0
  31. Rank #31

    claude-opus-4-5-20251101

    Anthropic

    Exact model version not yet confirmed

    Votes cast
    70,976
    License type
    Closed / proprietaryProprietary
    User preference score1469Based on anonymous head-to-head comparisons
    Likely score range1,465.91,472.3
  32. Rank #32

    ernie-5.1

    Baidu

    Exact model version not yet confirmed

    Votes cast
    37,418
    License type
    Closed / proprietaryProprietary
    User preference score1468Based on anonymous head-to-head comparisons
    Likely score range1,463.11,472.4
  33. Rank #33

    Official version confirmed

    Votes cast
    7,833
    License type
    Closed / proprietaryProprietary
    User preference score1468Based on anonymous head-to-head comparisons
    Likely score range1,460.41,475.0
    Review match evidence
  34. Rank #34

    mimo-v2.5-pro

    Xiaomi

    Exact model version not yet confirmed

    Votes cast
    42,195
    License type
    MIT
    User preference score1467Based on anonymous head-to-head comparisons
    Likely score range1,462.41,471.3
  35. Rank #35

    gpt-5.4

    OpenAI

    Exact model version not yet confirmed

    Votes cast
    62,036
    License type
    Closed / proprietaryProprietary
    User preference score1466Based on anonymous head-to-head comparisons
    Likely score range1,462.41,470.1
  36. Rank #36

    grok-4.1-thinking

    xAI

    Exact model version not yet confirmed

    Votes cast
    65,461
    License type
    Closed / proprietaryProprietary
    User preference score1466Based on anonymous head-to-head comparisons
    Likely score range1,462.71,469.1
  37. Rank #37

    qwen3.5-max-preview

    Alibaba

    Exact model version not yet confirmed

    Votes cast
    21,479
    License type
    Closed / proprietaryProprietary
    User preference score1465Based on anonymous head-to-head comparisons
    Likely score range1,459.81,469.8
  38. Rank #38

    Official version confirmed

    Model ID
    claude-sonnet-5
    Votes cast
    13,521
    License type
    Closed / proprietaryProprietary
    User preference score1461Based on anonymous head-to-head comparisons
    Likely score range1,454.91,467.4
    Review match evidenceModel version confirmed, but site comparison cannot yet preserve this ranking configuration
  39. Rank #39

    kimi-k2.6

    Moonshot AI

    Official version confirmed

    Votes cast
    37,686
    License type
    Modified MIT
    User preference score1461Based on anonymous head-to-head comparisons
    Likely score range1,456.41,465.5
    Review match evidence
  40. Rank #40

    qwen3.6-max-preview

    Alibaba

    Exact model version not yet confirmed

    Votes cast
    5,188
    License type
    Closed / proprietaryProprietary
    User preference score1460Based on anonymous head-to-head comparisons
    Likely score range1,451.71,468.5
Showing 40 of 378

Ranking basis

How to read this research edition

  • Scores stay on their original publisher scales. They are not combined into a composite score.

  • Missing data does not mean zero. It means the publisher did not report that metric for this snapshot.

  • When a source reports model settings, the position applies to that exact setup and source release—not every task, region, or deployment.

  • No provider can pay for placement. OpenGPT links directly to the evidence used.

Ranking basis

Data and sources

Each view preserves the source’s metric, release, and methodology. Follow the links to audit the underlying work.
AR

LMArena

The complete latest style-controlled overall ranking: all 378 source-listed model entries, based on users comparing two model answers side by side.

Leaderboard publication date
Jul 21, 2026
Data retrieved
Jul 24, 2026
License
CC BY 4.0

Original rankingScoring method

AA

Artificial Analysis

Nine evaluations across agentic work, coding, general knowledge, and scientific reasoning.

Source snapshot
Intelligence Index v4.1 · continuously updated
Data retrieved
Jul 24, 2026

Original rankingScoring method

LB

LiveBench

Twenty-three objective tasks across seven categories; questions refresh every six months.

Source snapshot
LiveBench-2026-06-25
Data retrieved
Jul 24, 2026

Original rankingScoring method

Method

A predictable publishing rhythm

OpenGPT will separate changing signals from frozen editions so past decisions remain auditable.
  1. 01

    Weekly watch

    Review new model and benchmark releases; flag material changes without silently rewriting this edition.

  2. 02

    Monthly frozen edition

    Publish a dated snapshot with exact configurations, source releases, and corrections.

  3. 03

    Quarterly method review

    Review categories, source quality, conflicts, and inclusion rules before changing the method.

Published

Past snapshots

August 2026August 2026

Scheduled next edition. It will be published only after source capture and review.