Data edition
2026-09Published ·
Automated decision scope
20 in automated verification / 91 catalog modelsData cutoff · 2026-09-09
Lifecycle coverage
56 lifecycle states documented / 91 · 20 in automation scopeData cutoff · 2026-09-09
OFFICIAL MODEL REFERENCE

AI model directory

Browse exact model IDs listed in provider documentation. Only versions reliably linked to comparable public data enter the rankings.

Model series
13
Official versions
91
Public ranking available
0
Awaiting comparable data
91
OFFICIAL MODEL REFERENCE

Official model IDs

Every tracked version remains visible. Ranked versions identify their data sources; other versions explain why no ranking is shown.

13 series · 91 versionsSubmit model
Deployment
Ranking data
13 / 13 series · 91 versions

Showing 2 of 13 model series

01 · 9 versions

GPT

A broad proprietary model family with a mature API and tool ecosystem.

Everyday workSolve problems or codeUse current information or tools
Official source

Compare versions

3 of 9 versions shown

GPT · Compare versions
Select for comparisonModel / versionBest fitPublic ranking dataOfficial access
gpt-6-astraHighest capability
Not yet rankedNo comparable public result is available for this exact version in the current edition.
gpt-5.5Deep reasoning
Not yet rankedNo comparable public result is available for this exact version in the current edition.
gpt-5.6-solHighest capability
Not yet rankedNo comparable public result is available for this exact version in the current edition.

Strengths

Covers general, coding, vision, and tool-based workflows.

Limitations

Closed weights and provider-controlled deployment.

02 · 8 versions

Claude

A proprietary family positioned for analysis, writing, coding, and long-context work.

Read, analyze, or writeRead long documentsSolve problems or code
Official source

Compare versions

3 of 8 versions shown

Claude · Compare versions
Select for comparisonModel / versionBest fitPublic ranking dataOfficial access
claude-fable-5-1Highest capability
Not yet rankedNo comparable public result is available for this exact version in the current edition.
claude-opus-5Quality and cost balance
Not yet rankedNo comparable public result is available for this exact version in the current edition.
claude-fable-5Highest capability
Not yet rankedNo comparable public result is available for this exact version in the current edition.

Strengths

Well suited to document analysis and sustained written work.

Limitations

Closed weights and API-dependent access.

Showing 2 of 13 model series
PUBLIC SOURCE INDEX

Complete source-listed model index

All 429 published records from 2 public sources, grouped into 423 unique source labels. Inclusion is not a recommendation; identity mappings are shown separately.

429 / 423
Published records
429
Unique source labels
423
Mapped to official IDs
0
Unresolved identities
423
Public source
Identity status

Showing 12 of 423 source labels

LiveBenchUnresolved identity
Source model ID

Qwen 3.8 Flash Next

Alibaba

Cost per successful task · #3 · $0.0423Open weights · #5 · 76.19
LiveBenchUnresolved identity
Source model ID

Qwen 3.8 Max

Alibaba

Objective tasks · #10 · 78.46Open weights · #2 · 78.46
LiveBenchUnresolved identity
Source model ID

Qwen3.8 27B

Alibaba

Open weights · #7 · 75.27
LiveBenchUnresolved identity
Source model ID

Claude 5 Opus Thinking Max Effort

Anthropic

Objective tasks · #7 · 80.08
LiveBenchUnresolved identity
Source model ID

Claude Fable 5 Max Effort

Anthropic

Objective tasks · #2 · 82.97
LiveBenchUnresolved identity
Source model ID

Claude Fable 5.1 Max Effort

Anthropic

Objective tasks · #1 · 83.41
LiveBenchUnresolved identity
Source model ID

DeepSeek V4 Flash 0731

DeepSeek

Cost per successful task · #6 · $0.0596Open weights · #8 · 74.17
LiveBenchUnresolved identity
Source model ID

DeepSeek V4 Flash Vision Exp

DeepSeek

Cost per successful task · #5 · $0.0505Open weights · #4 · 76.76
LiveBenchUnresolved identity
Source model ID

DeepSeek V4 Pro 0813

DeepSeek

Cost per successful task · #4 · $0.0442Open weights · #3 · 77.44
LiveBenchUnresolved identity
Source model ID

Gemini 3.5 Flash-Lite High

Google

Cost per successful task · #9 · $0.0686
LiveBenchUnresolved identity
Source model ID

Gemini 3.7 Flash High

Google

Objective tasks · #9 · 78.83
LiveBenchUnresolved identity
Source model ID

Muse Spark 1.3 xHigh Effort

Meta

Objective tasks · #4 · 81.59
Showing 12 of 423 source labels