{"slug": "opus-5-is-currently-1-on-artificial-analysis-intelligence-leaderboard", "title": "Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard", "summary": "Claude Opus 5 (Adaptive Reasoning, Max Effort) leads the Artificial Analysis Intelligence Index with a score of 61, out of 170 models evaluated, according to Artificial Analysis. The top five models are Claude Opus 5 variants and GPT-5.6 Sol, while Mercury 2 is the fastest at 938.7 tokens per second and Gemma 3n E4B Instruct is the most affordable at $0.02 per 1M tokens.", "body_md": "# Comparison of Models: Intelligence, Performance & Price Analysis\n\n[Microevals Playground](/microevals)\n\n[FAQs.](/faq)\n\n#### Intelligence\n\n#### Output Speed (tokens/s)\n\n#### Latency (seconds)\n\n#### Price ($ per M tokens)\n\n#### Context Window\n\nHighlights\n\n## Intelligence\n\n### Artificial Analysis Intelligence Index\n\n### Artificial Analysis Intelligence Index by Open Weights / Proprietary\n\n### Intelligence Evaluations\n\nAgentic real-world work tasks, (Elo-500)/2000\n\nAgentic tool use\n\nAgentic coding & terminal use\n\nCoding\n\nReasoning & knowledge\n\nScientific reasoning\n\nPhysics reasoning\n\nKnowledge\n\n1 - hallucination rate\n\nLong context reasoning\n\nAgentic knowledge work, Elo\n\nAgentic SaaS workflows\n\nLegal agentic work, criterion pass rate\n\nAgentic business operations\n\nInstruction following\n\nLong-horizon agentic tasks\n\nKubernetes incident root-cause analysis\n\nVisual reasoning\n\n## AA-Briefcase\n\n### AA-Briefcase Elo\n\n## AA-Omniscience\n\n### AA-Omniscience Index\n\n## Openness\n\n### Artificial Analysis Openness Index: Score\n\n## Intelligence Index Comparisons\n\n### Intelligence Index vs. Cost per Intelligence Index Task\n\n## Token Use\n\n### Output Tokens per Intelligence Index Task\n\n## Price and Cost\n\n### Cost per Intelligence Index Task\n\n### Cost to Run Artificial Analysis Intelligence Index\n\n### Pricing: Cache Hit, Input, and Output\n\n## Context Window\n\n### Context Window\n\n## Speed\n\nMeasured by Output Speed (tokens per second)\n\n### Output Speed\n\n### Time per Intelligence Index Task\n\n## Latency\n\nMeasured by Time (seconds) to First Token\n\n### Latency: Time To First Answer Token\n\n## End-to-End Response Time\n\nSeconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed\n\n### End-to-End Response Time\n\n## Model Size (Open Weights Models Only)\n\n### Model Size: Total and Active Parameters\n\n## Frequently Asked Questions\n\nClaude Opus 5 (Adaptive Reasoning, Max Effort) currently leads the Artificial Analysis Intelligence Index with a score of 61, out of 170 models evaluated.\n\nThe top AI models by Intelligence Index are: 1. Claude Opus 5 (Adaptive Reasoning, Max Effort) (61), 2. Claude Opus 5 (Adaptive Reasoning, Xhigh Effort) (60), 3. Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) (60), 4. GPT-5.6 Sol (max) (59), and 5. Claude Opus 5 (Adaptive Reasoning, High Effort) (59).\n\nMercury 2 is the fastest at 938.7 tokens per second, followed by HyperNova 60B 2605 (436.2 t/s) and Granite 4.0 H Small (431.0 t/s).\n\nGemma 3n E4B Instruct is the most affordable at $0.02 per 1M tokens (blended), followed by Nova Micro ($0.03) and Sarvam 30B (high) ($0.03).\n\nGemini 2.5 Flash-Lite (Non-reasoning) has the lowest time to first token at 0.35s, followed by Command A+ (0.41s) and Gemini 2.5 Flash (Non-reasoning) (0.52s).\n\nGLM-5.2 (max) is the highest-ranked open weights model with an Intelligence Index score of 51. There are 94 open weights models out of 170 total evaluated.\n\nThe top open weights AI models by Intelligence Index are: 1. GLM-5.2 (max) (51), 2. MiniMax-M3 (44), and 3. DeepSeek V4 Pro (Reasoning, Max Effort) (44).\n\nClaude Opus 5 (Adaptive Reasoning, Max Effort) leads among 126 reasoning models with an Intelligence Index score of 61. Reasoning models use extended thinking to work through complex problems before providing answers.\n\nModels are compared across multiple dimensions including intelligence (quality), pricing, output speed (tokens per second), latency (time to first token), end-to-end response time, and context window size. Performance metrics are measured directly using standardized prompts across 586 models.\n\nClick on any model name or row in the charts to view its dedicated page with detailed metrics and direct comparisons against similar models. You can also use the model selector to customize which models appear in each chart. [View the leaderboard](/leaderboards/models)", "url": "https://wpnews.pro/news/opus-5-is-currently-1-on-artificial-analysis-intelligence-leaderboard", "canonical_source": "https://artificialanalysis.ai/models", "published_at": "2026-07-24 19:45:10+00:00", "updated_at": "2026-07-24 22:23:45.820650+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-products", "ai-research"], "entities": ["Claude Opus 5", "Artificial Analysis", "GPT-5.6 Sol", "Mercury 2", "Gemma 3n E4B Instruct", "GLM-5.2", "DeepSeek V4 Pro", "Gemini 2.5 Flash-Lite"], "alternates": {"html": "https://wpnews.pro/news/opus-5-is-currently-1-on-artificial-analysis-intelligence-leaderboard", "markdown": "https://wpnews.pro/news/opus-5-is-currently-1-on-artificial-analysis-intelligence-leaderboard.md", "text": "https://wpnews.pro/news/opus-5-is-currently-1-on-artificial-analysis-intelligence-leaderboard.txt", "jsonld": "https://wpnews.pro/news/opus-5-is-currently-1-on-artificial-analysis-intelligence-leaderboard.jsonld"}}