# Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard

> Source: <https://artificialanalysis.ai/models>
> Published: 2026-07-24 19:45:10+00:00

# Comparison of Models: Intelligence, Performance & Price Analysis

[Microevals Playground](/microevals)

[FAQs.](/faq)

#### Intelligence

#### Output Speed (tokens/s)

#### Latency (seconds)

#### Price ($ per M tokens)

#### Context Window

Highlights

## Intelligence

### Artificial Analysis Intelligence Index

### Artificial Analysis Intelligence Index by Open Weights / Proprietary

### Intelligence Evaluations

Agentic real-world work tasks, (Elo-500)/2000

Agentic tool use

Agentic coding & terminal use

Coding

Reasoning & knowledge

Scientific reasoning

Physics reasoning

Knowledge

1 - hallucination rate

Long context reasoning

Agentic knowledge work, Elo

Agentic SaaS workflows

Legal agentic work, criterion pass rate

Agentic business operations

Instruction following

Long-horizon agentic tasks

Kubernetes incident root-cause analysis

Visual reasoning

## AA-Briefcase

### AA-Briefcase Elo

## AA-Omniscience

### AA-Omniscience Index

## Openness

### Artificial Analysis Openness Index: Score

## Intelligence Index Comparisons

### Intelligence Index vs. Cost per Intelligence Index Task

## Token Use

### Output Tokens per Intelligence Index Task

## Price and Cost

### Cost per Intelligence Index Task

### Cost to Run Artificial Analysis Intelligence Index

### Pricing: Cache Hit, Input, and Output

## Context Window

### Context Window

## Speed

Measured by Output Speed (tokens per second)

### Output Speed

### Time per Intelligence Index Task

## Latency

Measured by Time (seconds) to First Token

### Latency: Time To First Answer Token

## End-to-End Response Time

Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed

### End-to-End Response Time

## Model Size (Open Weights Models Only)

### Model Size: Total and Active Parameters

## Frequently Asked Questions

Claude Opus 5 (Adaptive Reasoning, Max Effort) currently leads the Artificial Analysis Intelligence Index with a score of 61, out of 170 models evaluated.

The top AI models by Intelligence Index are: 1. Claude Opus 5 (Adaptive Reasoning, Max Effort) (61), 2. Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) (60), 3. Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) (60), 4. GPT-5.6 Sol (max) (59), and 5. Claude Opus 5 (Adaptive Reasoning, High Effort) (59).

Mercury 2 is the fastest at 938.7 tokens per second, followed by HyperNova 60B 2605 (436.2 t/s) and Granite 4.0 H Small (431.0 t/s).

Gemma 3n E4B Instruct is the most affordable at $0.02 per 1M tokens (blended), followed by Nova Micro ($0.03) and Sarvam 30B (high) ($0.03).

Gemini 2.5 Flash-Lite (Non-reasoning) has the lowest time to first token at 0.35s, followed by Command A+ (0.41s) and Gemini 2.5 Flash (Non-reasoning) (0.52s).

GLM-5.2 (max) is the highest-ranked open weights model with an Intelligence Index score of 51. There are 94 open weights models out of 170 total evaluated.

The top open weights AI models by Intelligence Index are: 1. GLM-5.2 (max) (51), 2. MiniMax-M3 (44), and 3. DeepSeek V4 Pro (Reasoning, Max Effort) (44).

Claude Opus 5 (Adaptive Reasoning, Max Effort) leads among 126 reasoning models with an Intelligence Index score of 61. Reasoning models use extended thinking to work through complex problems before providing answers.

Models are compared across multiple dimensions including intelligence (quality), pricing, output speed (tokens per second), latency (time to first token), end-to-end response time, and context window size. Performance metrics are measured directly using standardized prompts across 586 models.

Click on any model name or row in the charts to view its dedicated page with detailed metrics and direct comparisons against similar models. You can also use the model selector to customize which models appear in each chart. [View the leaderboard](/leaderboards/models)
