cd/entity/Artificial Analysis· home entities Artificial Analysis
grep -l @artificial analysis /news/*.json | wc -l → 246

Artificial Analysis

mentions 246 type Person page 11/13 feed RSS

// recent coverage 246 mentions

00:00
2026-06-25
runagentrun.co.uk
large-language-models

Gemma 4 outpaces Qwen 3.6 on code review

Google's Gemma 4 31B outperforms Alibaba's Qwen 3.6 27B on agentic code review tasks, finishing faster due to superior Multi-Token Prediction (MTP) design, according to benchmarks and field reports. W…

21:53
2026-06-24
baseten.co
large-language-models

How we built the fastest API for GLM-5.2

Baseten has built the world's fastest API for GLM-5.2, achieving over 280 tokens per second as measured by Artificial Analysis. The performance is driven by optimizations including an updated inferenc…

12:00
2026-06-23
telnyx.com
large-language-models

GLM-5.2 is now available on Telnyx Inference

Telnyx has added Z.ai's GLM-5.2 open-weight model to its Inference platform, hosted on owned B300 GPUs. The model leads Artificial Analysis with an Intelligence Index of 51, outperforming competitors,…

00:00
2026-06-22
simpletechguides.com
large-language-models

Kimi K2.7 vs GLM 5.2 for Coding

Open-weight MoE models Kimi K2.7 Code and GLM 5.2, released in June 2026, target developers and agentic coding workflows but differ in design. GLM 5.2 leads with a 1M-token context window and lower ha…

10:12
2026-06-20
byteiota.com
large-language-models

GLM-5.2 Hallucinates 3x Less Than GPT-5.5 — Open Weight Wins

Z.ai released GLM-5.2 under an MIT license on June 16, and benchmarks show the open-weight model hallucinates at 28% versus GPT-5.5's 86% on the AA-Omniscience benchmark, a 3x reliability gap. GLM-5.2…

00:00
2026-06-20
runagentrun.co.uk
artificial-intelligence

AA-Briefcase: a tougher test for agents

Artificial Analysis released AA-Briefcase, a new agentic benchmark for long-horizon knowledge work, on June 18, 2026. Claude Fable 5 leads the leaderboard with 1587 Elo at $31 per task, while open-wei…

23:09
2026-06-19
mukulsingh105.github.io
artificial-intelligence

Knowledge workers don't need frontier models

A new architecture using a nano-model router that dispatches knowledge-worker tasks to either a frontier model or a small, cheap model achieves near-frontier quality at a fraction of the cost, ranking…

16:11
2026-06-19
arrowtsx.dev
large-language-models

GPT-5.5 hallucinates 3x more than MIT-licensed GLM-5.2

New benchmarks reveal that larger AI models like GPT-5.5 and DeepSeek V4 Pro hallucinate significantly more than smaller open-weight models, with GLM-5.2 achieving a 28% hallucination rate compared to…

15:30
2026-06-18
yaroslavvb.github.io
artificial-intelligence

How far behind is open-source AI?

The gap between open-source and proprietary AI models has narrowed from nearly 10 months in December 2024 to just 2–3.5 months as of early 2025, according to an analysis of the open-source Pareto fron…

11:01
2026-06-18
lesswrong.com
large-language-models

How far do open weights trail the frontier?

A new analysis using Epoch's ECI metric shows that open-weight AI models continue to trail closed models on the frontier, with the gap persisting over time. The analysis, based on item response theory…

10:16
2026-06-18
dev.to
large-language-models

What GLM-5.2 Changes for Long-Horizon Coding

Zhipu AI released GLM-5.2, a large language model with a 1M-token context window, flexible effort levels, and an MIT license, targeting long-horizon coding tasks. The model introduces IndexShare, an a…

09:16
2026-06-18
sebastianraschka.com
large-language-models

GLM-5.2 and IndexShare for Long-Context Sparse Attention

Z.ai released GLM-5.2, an open-weight model that the author calls the best open-weight model available. The model introduces IndexShare, a cross-layer reuse trick for DeepSeek Sparse Attention that re…

07:57
2026-06-18
dev.to
large-language-models

Nemotron 3 Ultra went live June 4. Here's the call that works.

NVIDIA released Nemotron 3 Ultra on June 4, 2026, a 550-billion-parameter open-weights model that achieves the highest intelligence score among US open models. The model uses a hybrid Mamba-Transforme…

23:58
2026-06-17
simonwillison.net
large-language-models

GLM-5.2 is probably the most powerful text-only open weights LLM

Chinese AI lab Z.ai released GLM-5.2, a 753B-parameter open-weights text-only LLM with a 1 million token context window, under an MIT license. The model leads the Artificial Analysis Intelligence Inde…

20:24
2026-06-15
aws.amazon.com
artificial-intelligence

Introducing Gemma 4 models on Amazon Bedrock

Amazon Bedrock announced the availability of Gemma 4 models, a family of open-weight AI models from Google DeepMind, including dense and mixture-of-experts variants with built-in reasoning, function c…

← prev page 11 / 13 next →
// co-occurs with top 8 entities
// topics top 6 topics