Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard Claude Opus 5 (Adaptive Reasoning, Max Effort) leads the Artificial Analysis Intelligence Index with a score of 61, out of 170 models evaluated, according to Artificial Analysis. The top five models are Claude Opus 5 variants and GPT-5.6 Sol, while Mercury 2 is the fastest at 938.7 tokens per second and Gemma 3n E4B Instruct is the most affordable at $0.02 per 1M tokens. Comparison of Models: Intelligence, Performance & Price Analysis Microevals Playground /microevals FAQs. /faq Intelligence Output Speed tokens/s Latency seconds Price $ per M tokens Context Window Highlights Intelligence Artificial Analysis Intelligence Index Artificial Analysis Intelligence Index by Open Weights / Proprietary Intelligence Evaluations Agentic real-world work tasks, Elo-500 /2000 Agentic tool use Agentic coding & terminal use Coding Reasoning & knowledge Scientific reasoning Physics reasoning Knowledge 1 - hallucination rate Long context reasoning Agentic knowledge work, Elo Agentic SaaS workflows Legal agentic work, criterion pass rate Agentic business operations Instruction following Long-horizon agentic tasks Kubernetes incident root-cause analysis Visual reasoning AA-Briefcase AA-Briefcase Elo AA-Omniscience AA-Omniscience Index Openness Artificial Analysis Openness Index: Score Intelligence Index Comparisons Intelligence Index vs. Cost per Intelligence Index Task Token Use Output Tokens per Intelligence Index Task Price and Cost Cost per Intelligence Index Task Cost to Run Artificial Analysis Intelligence Index Pricing: Cache Hit, Input, and Output Context Window Context Window Speed Measured by Output Speed tokens per second Output Speed Time per Intelligence Index Task Latency Measured by Time seconds to First Token Latency: Time To First Answer Token End-to-End Response Time Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed End-to-End Response Time Model Size Open Weights Models Only Model Size: Total and Active Parameters Frequently Asked Questions Claude Opus 5 Adaptive Reasoning, Max Effort currently leads the Artificial Analysis Intelligence Index with a score of 61, out of 170 models evaluated. The top AI models by Intelligence Index are: 1. Claude Opus 5 Adaptive Reasoning, Max Effort 61 , 2. Claude Opus 5 Adaptive Reasoning, Xhigh Effort 60 , 3. Claude Fable 5 Adaptive Reasoning, Max Effort, Opus 4.8 Fallback 60 , 4. GPT-5.6 Sol max 59 , and 5. Claude Opus 5 Adaptive Reasoning, High Effort 59 . Mercury 2 is the fastest at 938.7 tokens per second, followed by HyperNova 60B 2605 436.2 t/s and Granite 4.0 H Small 431.0 t/s . Gemma 3n E4B Instruct is the most affordable at $0.02 per 1M tokens blended , followed by Nova Micro $0.03 and Sarvam 30B high $0.03 . Gemini 2.5 Flash-Lite Non-reasoning has the lowest time to first token at 0.35s, followed by Command A+ 0.41s and Gemini 2.5 Flash Non-reasoning 0.52s . GLM-5.2 max is the highest-ranked open weights model with an Intelligence Index score of 51. There are 94 open weights models out of 170 total evaluated. The top open weights AI models by Intelligence Index are: 1. GLM-5.2 max 51 , 2. MiniMax-M3 44 , and 3. DeepSeek V4 Pro Reasoning, Max Effort 44 . Claude Opus 5 Adaptive Reasoning, Max Effort leads among 126 reasoning models with an Intelligence Index score of 61. Reasoning models use extended thinking to work through complex problems before providing answers. Models are compared across multiple dimensions including intelligence quality , pricing, output speed tokens per second , latency time to first token , end-to-end response time, and context window size. Performance metrics are measured directly using standardized prompts across 586 models. Click on any model name or row in the charts to view its dedicated page with detailed metrics and direct comparisons against similar models. You can also use the model selector to customize which models appear in each chart. View the leaderboard /leaderboards/models