cd/entity/Artificial Analysis· home entities Artificial Analysis
grep -l @artificial analysis /news/*.json | wc -l → 246

Artificial Analysis

mentions 246 type Person page 8/13 feed RSS

// recent coverage 246 mentions

23:02
2026-07-16
artificialanalysis.ai
artificial-intelligence

Kimi K3 beats GPT 5.6 Sol in agentic knowledge work

Kimi K3 from Moonshot AI has surpassed OpenAI's GPT-5.6 Sol in agentic knowledge work, achieving an Elo rating of 1547 compared to GPT-5.6 Sol's 1495 on the AA-Briefcase benchmark, which evaluates mod…

22:28
2026-07-16
tokenstead.ai
artificial-intelligence

Kimi K3

Moonshot AI released Kimi K3, a 2.8-trillion-parameter mixture-of-experts model with 16 active experts per token, achieving the top score of 1679 on the Artificial Analysis webdev arena ahead of Claud…

20:09
2026-07-16
artificialanalysis.ai
artificial-intelligence

Kimi K3 Intelligence, Performance and Price Analysis

Kimi K3, a reasoning model released on July 16, 2026, by Kimi, scores 57 on the Artificial Analysis Intelligence Index, well above the average of 30 among comparable models, but is slower than average…

19:29
2026-07-16
aws.amazon.com
artificial-intelligence

Introducing Grok on Amazon Bedrock

XAI's Grok 4.3 is now generally available on Amazon Bedrock, offering configurable reasoning effort, a 1 million token context window, and tool use for building agents. The model runs on Mantle, Amazo…

15:04
2026-07-16
platform.kimi.ai
artificial-intelligence

Introducing Kimi K3

Kimi released Kimi K3, its most capable model with 2.8 trillion parameters, built on Kimi Delta Attention and Attention Residuals, offering native visual understanding and a 1M-token context window. I…

14:36
2026-07-16
artificialanalysis.ai
artificial-intelligence

Inkling Benchmark Results

Thinking Machines has released Inkling, a 975B-parameter open weights model with 41B active parameters, debuting at 41 on the Artificial Analysis Intelligence Index and becoming the leading open weigh…

02:53
2026-07-15
artificialanalysis.ai
large-language-models

GPT-5.6 Sol, Terra, Luna compare on intelligence vs. cost

GPT-5.6 Sol and Luna outperform Terra at every point on the Intelligence vs Cost per Task chart, with Luna emerging as a particularly cost-efficient model, according to the Artificial Analysis Intelli…

07:54
2026-07-14
techstrong.ai
artificial-intelligence

You Can Keep the Benchmarks. I’ll Take the Test Drive

OpenAI's GPT-5.6 Sol and Anthropic's Claude Fable 5 represent a substantial advance in AI, handling complex, multi-step work with less human guidance, according to a hands-on evaluation by a senior ne…

06:09
2026-07-14
artificialanalysis.ai
artificial-intelligence

Harvey LAB-AA: evaluating AI agents on real-world legal work

Harvey LAB-AA, a new benchmark from Artificial Analysis evaluating AI agents on real-world legal work across 24 practice areas, shows Claude Fable 5 (max, with Opus 4.8 fallback) leading with a 14.2% …

00:00
2026-07-13
tomtunguz.com
artificial-intelligence

The AI Colander

AI models retain between high single digits and 40% of customers after five months, with the stickiest foundational cohorts near the top of that range, according to a study by OpenRouter and a16z. The…

← prev page 8 / 13 next →
// co-occurs with top 8 entities
// topics top 6 topics